Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.06565.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:51:34.192526Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-28T23:02:46.504962Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 5bb5e92d-9583-499e-abeb-39b9c576a79d · inbound
Challenges in Trustworthy Human Evaluation of Chatbots MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1fd342f-1825-465d-aacd-248e310052f3 · inbound
C$^2$LEVA: Toward Comprehensive and Contamination-Free Language Model Evaluation MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082c3b24-4b68-4219-8160-83facebafea4 · inbound
Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be3cb942-fb86-4ae2-9b9c-c642367a9f62 · inbound
Enhancing LLMs for Governance with Human Oversight: Evaluating and Aligning LLMs on Expert Classification of Climate Misinformation for Detecting False or Misleading Claims about Climate Change MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff803e47-8d9f-4f93-840e-4dd68681c9d5 · inbound
Tuning LLM Judge Design Decisions for 1/1000 of the Cost MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6358db70-4ec3-4681-aa24-5c684c816f7e · inbound
How to Select Datapoints for Efficient Human Evaluation of NLG Models? MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c174b93-ad88-4921-ba29-320fe7494e34 · inbound
WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7120ae5c-e864-4839-b264-e5eee69331c8 · inbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31acb379-afae-4f1f-953f-aa1742758e0c · inbound
When AI Co-Scientists Fail: SPOT-a Benchmark for Automated Verification of Scientific Research MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c168072-8d54-4881-b6d9-30118242fd5f · inbound
Decentralized Arena: Towards Democratic and Scalable Automatic Evaluation of Language Models MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77a84dd6-360f-48fe-af96-b2995dd82fd5 · inbound
Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d8b17e7-3c82-493d-9fb2-3ca2bb6048a2 · inbound
Mellum2 Technical Report MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.