Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:31:00.881052Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2506.05062.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:31:00.881052Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 33d6d86c-72c1-42ed-a735-2c4df29cb826 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 39b7c9d9-06ab-4d2e-880d-29cf93d08887 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Most arguments in this speech support the topic
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a864ccc4-1335-4394-a70e-f46e328212bf · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Pipeline-set-1
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b187ff4-7b09-43ef-b547-33a2ccbcb096 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 074e1a82-a236-4b20-a8dc-06cf6f5cd173 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6f4aa0ec-40a9-488b-ac06-e7b466a42749 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab511778-c2a4-42c0-b5da-9974dc366b36 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation The argument for reform is strong
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b86c1e15-fc16-49e2-a1b0-13b486042ea9 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation In general, results for the CoT prompt seem to be more challenging to parse
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 02956fd4-41e3-4089-85f0-214d724bc912 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e2db4569-ce24-402a-bb11-70f45758816c · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 575f073d-bfac-4c1e-a1c3-2fcc5205040a · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a17ae3ce-5541-47ac-8fc1-431e287fc676 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation This speech is a good opening speech for supporting the topic
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb420f3d-ed82-4e07-ae54-8b78c3c98301 · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation InProceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pages 7073–7086, Online
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d791c09-863a-4a28-9d48-9f061c101aeb · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation ChatGPT as a Factual Inconsistency Evaluator for Text Summarization
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc470726-7043-45dd-b61e-e400e80946ce · outbound
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation This speech is a good opening speech for supporting the topic
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.