Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T10:02:42.228482Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2607.16239.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T10:02:42.228482Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cc0b5e90-ca7b-4f89-abdb-5a5af7d5d4e0 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges and Tanno, Ryutaro and Schwaighofer, Anton and Tezcan, Kerem C
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88db8488-36db-4107-9139-1e0ef6d2ec70 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges CROWDLAB: Supervised learning to infer consensus labels and quality scores for data with multiple annotators
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a826fa97-6006-4213-9e99-bf46f3874c23 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Active Label Correction for Semantic Segmentation with Foundation Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 146248a4-36f8-49dc-85f0-ed55d9f7b979 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Re-labeling ImageNet: from Single to Multi-Labels, from Global to Localized Labels
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5d3000d0-13a7-4bfb-8455-5a40d5cdd5d4 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Pervasive Label Errors in Test Sets Destabilize Machine Learning Benchmarks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb514b99-9dfb-41ae-ac39-b5140524b364 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Confident Learning: Estimating Uncertainty in Dataset Labels
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86b43fe6-9d1e-44be-aa2c-844a416ae05d · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges International Journal of Human-Computer Studies , volume=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68f15005-44a2-4874-804f-f6a3c4668752 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0947870-4e18-42da-a976-01d0b9beac98 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges 2016 , pages =
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 805558bc-e207-4f83-8057-4f4c8011a8a9 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Applied Statistics , author =
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d06bfe2-70b1-4e8a-8367-61f2a05bfdb3 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b2a61fe-472f-44f8-84cd-293b4924161f · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44171479-06ac-4ad4-ab81-5e0ff22c3905 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49956311-0aee-4a9e-9b69-bde3928a114c · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges 2002 , publisher=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eea6e29f-18e6-4e30-9230-d393a2c765aa · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges arXiv preprint arXiv:2507.01372 , year=
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ae0b878-47b0-4fa7-9dee-d617f8312307 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges arXiv preprint arXiv:2511.08991 , year=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25edf008-2d97-419b-bc87-9ad7750d29e1 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Deep Think with Confidence
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 180fef72-91aa-47ab-a012-5dc417b0be68 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Assessing Writing , volume=
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18997e56-ebde-4f3f-b138-1f0cb174080a · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Transactions of the Association for Computational Linguistics , volume=
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d30b06c9-9b1a-401b-b2b0-71c52a21e0bc · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Proceedings of the IEEE/CVF international conference on computer vision , pages=
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1e7c997-729b-4c6a-b6c5-99f0958b1d75 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges The Annals of Applied Statistics , volume=
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b46fb76-60b6-46fe-8524-93d8ee668a64 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges The Econometrics Journal , volume=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23f5c53a-263d-4bb1-b4bb-7fe6dce984de · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Science , volume=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43681a52-226c-48a9-93ee-c1ed77290899 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges and Duchi, John C
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23442581-0629-44f7-8d24-7ca5e36ffe3c · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Proceedings of the National Academy of Sciences , volume=
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22d79700-0cdf-41f3-9706-e4f5ebbd2704 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges 1992 , publisher=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35910408-382e-4de7-a997-10550edc747c · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Biometrika , volume=
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 390dfa8b-5c5e-411e-ac97-de27f2cb148a · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges The Annals of Statistics , volume=
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69c94ca9-91b8-427f-b2ca-cdef1f05ae4a · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Journal of the Royal Statistical Society: Series B , volume=
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf7d139b-f6f9-4d58-9ed6-ed55e71747b5 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges and Zhang, Hao and Gonzalez, Joseph E
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69cf0b29-3197-48c7-982f-a2e93031c653 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84348907-2062-41d7-a97f-73f1e75be3ad · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Large Language Models are not Fair Evaluators
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 009051d1-630f-48b8-b4f9-d3e81c4f086f · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 043dcf38-241a-44bb-9433-c16c5a270922 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges A Survey on
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2c319f9-468e-435f-a625-85c7e7285589 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97f940c-2b34-4c41-85fd-99f6f44f02f8 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Proceedings of the 41st International Conference on Machine Learning , year=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a03f5688-6074-49aa-8dc7-4053e3456085 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , year=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9aa15a8-84c8-4960-bec9-bee4cc7f4f75 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d70c31c-173d-4a66-97b1-fbff00907b91 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges International Conference on Learning Representations , year=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 199b660b-6ff6-45d6-a181-5536b142840e · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Replacing Judges with Juries: Evaluating LLM Generations with a Panel of Diverse Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba43dac-c646-460f-b004-3b9fb3f5d0a3 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Enhancing
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdd547c2-ebc6-452f-97ed-0833cffc93f8 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59775538-2086-444f-b486-9a5196eb6948 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges International Conference on Learning Representations , year=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da6dcea5-c533-4d7a-bd14-53e25c6e8103 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af2e6d10-3034-4383-bf28-324a9ee37afd · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56667f6b-9cd5-43b1-9e61-0c3b7106d3b4 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Generative Judge for Evaluating Alignment
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d29f28c-c689-4eaf-b0bc-fbc72086e9f2 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges An Empirical Study of
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdda88e8-9cd7-4ac9-b8c0-4a668773beff · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 686ec06d-b12d-4a78-8e43-70fb67392f8e · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a34268e3-5feb-4ee4-ad66-b61e89733afe · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges The Computational Geometry Algorithms Library , subtitle =
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50eef0ae-6985-40d2-9628-3cefc3253b21 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 784d6f4c-fe0c-4d32-9625-955b616b58b4 · outbound
BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.