Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:58:27.657384Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 0 inbound Pith citation observations for arXiv:2506.03785.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:58:27.657384Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
11 of 11 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 331d1f76-3d4d-4d0a-9293-1f544e829147 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 647f39a9-061e-4496-b608-7e577fe028fd · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 85d3b424-5874-4216-83a8-654b2f96e3ad · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fceb344c-b0ce-4cbc-aa62-dcc7fff13356 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df6d0917-afb7-454d-b77f-f2a6d1a7b858 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3abebc9c-4799-4c0d-9afd-db3fab180062 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 333b0b5b-1c0a-4de9-9522-bc719485b433 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Large Language Models are Biased Because They Are Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d54c5bac-9273-48d6-b8c2-88b91f717d13 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Is ChatGPT a Good NLG Evaluator? A Preliminary Study
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9f5cabf-7b10-4ae3-8b80-50c79fa0b58e · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dc644c4-dcf5-48f5-ad19-b45acd21457d · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons online" 'onlinestring :=
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604b1408-b9d0-4fb3-970b-6e7158486d25 · outbound
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons write newline
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.