Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2309.17012.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:03.842968Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-07T12:53:50.367532Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation f899883b-0a27-470b-8d38-f46471bf40bf · inbound
LLM Evaluators Recognize and Favor Their Own Generations Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab7b2c8e-08ff-4d0d-838d-9aed89e9baa7 · inbound
Better & Faster Large Language Models via Multi-token Prediction Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fc4e8733-be44-49d0-9fec-7d700a4c89f8 · inbound
From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a3abb5e-cfdd-4880-860c-93fc57e331bd · inbound
A Survey on LLM-as-a-Judge Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9e50b7f7-7e23-4f03-9e13-a8302fd0db5b · inbound
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 116
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f4d8d5a-931a-4bb6-81bc-60471a111347 · inbound
Beyond the Surface: Measuring Self-Preference in LLM Judgments Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c134385-792b-4c51-9ea6-155b1800f4e3 · inbound
AbsenceBench: Language Models Can't Tell What's Missing Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4568b5f1-b150-4c37-96a1-ad7f735d3388 · inbound
CRISP: Complex Reasoning with Interpretable Step-based Plans Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 602d4ffe-c342-4bef-9e26-f28960e02120 · inbound
Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44abe90e-85c8-43a4-9e08-15a1b0802d92 · inbound
Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1b445b0-1406-4459-b511-3d2c5ce62956 · inbound
Can You Trick the Grader? Adversarial Persuasion of LLM Judges Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e1e2dec-7704-4b8c-935e-200e8f2c817e · inbound
Effectively obtaining acoustic, visual and textual data from videos Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4a5ec78-d9dc-4d66-ba76-3f0446d240da · inbound
Testing chatbots on the creation of encoders for audio conditioned image generation Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 038f4fa5-de01-4194-bad5-b298963f42eb · inbound
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b59a23e2-396e-4e0c-a714-32830efcc736 · inbound
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ed94691-671b-45a0-9090-854487c94667 · inbound
Self-Preference Bias in Rubric-Based Evaluation of Large Language Models Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38c552b8-0b23-44f9-86b9-13366418ef4a · inbound
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2ffbd726-6c6a-476d-9f0e-873e74becc45 · inbound
Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e1a4976e-a1e9-4a7e-94f5-0d076211b651 · inbound
U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 66475279-6144-48ed-9d7e-c63011334597 · inbound
Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8bbb5c07-bfda-4855-9660-e75ed527d963 · inbound
Does Capability Transfer to Subjective Behavior -- and Would Our Instruments Tell Us? A Self-Evolving, Trust-by-Construction Evaluation Paradigm Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb9f1fd1-16a0-48c7-81b5-28a198b0db93 · inbound
Show, Don't TELL: Explainable AI-Generated Text Detection Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 50aad0d0-743e-43ca-9fad-4a0f00ca67f3 · inbound
Poller: Are LLMs Suitable for Evaluating the Poetry Understanding Task? Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1892a6d9-f0ed-461b-86c8-9c57c92a65c1 · inbound
LLM-as-a-Verifier: A General-Purpose Verification Framework Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c9b73bbf-03de-4ba3-b178-d538abc065d4 · inbound
LLM-as-a-Verifier: A General-Purpose Verification Framework Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7ec13a-8c12-44aa-9f13-a015ba53d632 · inbound
AMT-X: Phase-Structured Multi-Turn Red-Teaming with Checklist-Gated Evaluation Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cb7b65c-3a5a-4d79-9294-95f4a01daa4d · inbound
Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48c6a736-6094-48ba-be94-2cf8e6e86dbc · inbound
What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces Benchmarking Cognitive Biases in Large Language Models as Evaluators
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.