Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:42:54.008877Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 1 inbound Pith citation observation for arXiv:2505.12301.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:42:54.008877Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:37:29.936416Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-12T14:37:31.118843Z
56 of 56 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee8d0a4e-5762-49da-9adb-0a050332e44c · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge A Survey of Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc1d6655-26d1-4fa8-afa6-fdf5f1276cbd · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Language models are few-shot learners
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0850222e-0e5e-46d7-99b8-b884493fa47c · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Chi, Quoc V
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b639d5e7-d5fd-4501-964a-2b414e3deef5 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Reflexion: language agents with verbal reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 58b73ae6-2475-4013-b0f6-eed74d0755be · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Reasoning with large language models, a survey, 2024
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 208bbd3c-88e7-4589-814c-f3236d9e26f6 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Evaluating and improving tool-augmented computation-intensive math reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 29bf1587-ca44-40a0-ab50-52eace44d2fb · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb482f53-2044-413d-8be0-02e4cd35faba · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge A Survey on LLM-as-a-Judge
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61e604f9-7db1-464a-8e06-fc72f8e6babf · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Lima: less is more for alignment
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ea7ad07c-3221-4bfe-939b-d728296c6e65 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Flask: Fine-grained language model evaluation based on alignment skill sets
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 10b3df07-1009-45f6-9880-a3111c8fc77a · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge JudgeLM: Fine-tuned Large Language Models are Scalable Judges
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3722f0e-b4f9-4c46-9d1f-ec3b004b4ecb · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Wildbench: Benchmarking llms with challenging tasks from real users in the wild
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8b1bdf3b-d4dd-41c8-9419-0d5a43efb2ca · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Judgment under uncertainty: Heuristics and biases: Biases in judgments reveal some heuristics of thinking under uncertainty.science, 185(4157): 1124–1131, 1974
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1844526f-7288-4cd0-ad9d-f48716519e63 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Beyond correlation: The impact of human uncertainty in measuring the effectiveness of automatic evaluation and LLM-as-a-judge
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9ea58b09-609f-4290-ba60-48db2a7d89a0 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Towards measuring the representation of subjective global opinions in language models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 31d60ee8-4b0c-4f96-aee7-a2bc855b02c9 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e05c6dd-6cec-4989-9e9f-c2dc8cca3da2 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ee778e-a63b-47b0-881b-5537dd2ee27d · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge On information and sufficiency.The annals of mathematical statistics, 22(1):79–86, 1951
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c95278a-da20-4fcd-aeb1-1c8084362869 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge ChatGPT as a Factual Inconsistency Evaluator for Text Summarization
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4d9ef74-2829-4fe6-b1a5-46560a40862a · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Human-like Summarization Evaluation with ChatGPT
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12e49271-933f-438e-bdcb-8598792475f7 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Distributional preference learning: Understanding and accounting for hidden context in RLHF
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6e3a5a1f-b1c7-4e01-9524-678cd4e21e0a · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Aligning crowd feedback via distributional preference reward modeling
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2d18b654-3c6c-4188-805d-813beb713306 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Quantile Regression for Distributional Reward Models in RLHF
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb120bcf-831f-4694-ab47-551b2b5e1ea2 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Improving llm-as-a-judge inference with the judgment distribution.arXiv preprint arXiv:2503.03064, 2025
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c676855-0897-4006-812f-82517e0b7e95 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Fre- quency domain adversarial training for robust volumetric medical segmentation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b6a5ca05-976f-496a-8ad2-2f2519ba947e · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Self-adaptive adversarial training for robust medical segmentation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9952b9e1-7039-47b4-8fd5-c3bba42f40f7 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Adversarial training methods for semi- supervised text classification
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 25ecb149-fd52-4328-99ec-9ac50de30490 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Adversarial examples for evaluating reading comprehension systems
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b447e46e-e085-4163-b911-ab3a95df6bcf · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge A survey of robust adversar- ial training in pattern recognition: Fundamental, theory, and methodologies.Pattern Recognition, 131:108889, 2022
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9816a2ac-53f0-4c75-80c9-7832ea5b99fa · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Explaining and Harnessing Adversarial Examples
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f324eef5-436b-418a-bcd9-62707e113d35 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Towards deep learning models resistant to adversarial attacks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f7d0343-465b-49f5-8083-1f94d241d367 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Summeval: Re-evaluating summarization evaluation.Transactions of the Association for Computational Linguistics, 9:391–409, 2021
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8bb4ff40-997a-45bd-a50c-4c0ab69f5ec6 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge G- eval: NLG evaluation using gpt-4 with better human alignment
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17a8750e-4674-45e6-a17f-62aec0bb37da · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Distilling the knowledge in a neural network
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fe2cdade-bc0a-443e-8ab9-b4d6e3ad5e74 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 899ab686-aeb3-4595-893c-3ff421b7ed1a · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation db8c5bf2-bc94-4442-ad18-8b2c09db98e5 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge CVXPY: A Python-embedded modeling language for convex optimization.Journal of Machine Learning Research, 17(83):1–5, 2016
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f73c8b9-1926-4eec-b196-7415d7557ca3 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Bowman, Gabor Angeli, Christopher Potts, and Christopher D
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f6262e9-c453-4f01-ae5d-9c53f5b5b39f · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge A broad-coverage challenge corpus for sentence understanding through inference
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 21cbffdc-17c0-41a7-a03c-cf6a6a5dd619 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Qwen2.5 Technical Report
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c64ef661-d4b3-44ac-9659-cc7ebb97a7a9 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Xing, Hao Zhang, Joseph E
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8831e236-8f31-4510-844f-7cd191e167e9 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Decoupled weight decay regularization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a82480fd-ac61-45cd-b51f-e4e79dfabcfa · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge The llama 3 herd of models,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 5dc8fa0a-d2fd-43a2-b618-a5dece0055f3 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Evaluating the logical reasoning ability of chatgpt and gpt-4, 2023
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e48f3eb6-a54b-44de-8b93-d7a3bebd124c · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fff391d-e52c-4df0-afda-fb35594b03c4 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Check if the summary covers the main topic and key points of the news article, and if it presents them in a clear and logical order
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 810b898d-c51c-45cf-8e27-d942aadf805a · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge 1", "2",
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 951a2391-fb1b-47d6-943e-f1751702ca33 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff94eb2-451e-4f0a-8ab9-d74b4bd94b6a · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Check if the summary contains any factual errors that are not supported by the article
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7ce1f34a-a823-4e95-ad30-590c8edcb838 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge 1", "2",
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dc0d288a-eb7b-45b0-a4d7-7eeeed7ab956 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 333e8bef-0d2e-4bca-aec6-b3ffeede76be · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e51b784f-a0ae-46d9-b4aa-37ec5642e390 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86fb61d0-fde4-4f06-baaf-a1da97443991 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge 1", "2",
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 777ca137-a3d5-400b-adb1-aa5f2f594b71 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge doi: 10.18653/v1/N18-1101
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0daeec2b-d25f-4ac8-a4a6-3aeff54c7f30 · outbound
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge The Llama 3 Herd of Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8159b61-b47f-461f-8137-aa869d430bdd · inbound
Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.