Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2410.17578.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:09.535726Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation aaaaff09-b013-4962-823d-53722368e49a · inbound
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 209
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a5a099c4-8d9f-457a-9cf1-4728329dce4f · inbound
Controlling Language Confusion in Multilingual LLMs MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f53474a-1f8d-47e8-b372-0baa0c5980b3 · inbound
IMPACT: Inflectional Morphology Probes Across Complex Typologies MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80d65daf-bf30-4c19-a3bb-6c5f5417cd2a · inbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55aa138a-6109-4068-afdf-a06fcc171e31 · inbound
Multilingual Prompt Localization for Agent-as-a-Judge: Language and Backbone Sensitivity in Requirement-Level Evaluation MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1ac2f856-344c-4bd4-b0fc-19e7e828ad40 · inbound
ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5eee0f55-d720-4fea-b851-e1af78d24fd6 · inbound
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f7425033-3a6d-401a-ad10-585a4367b12a · inbound
Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 466da1dc-1a78-458c-ac5b-b9df4a8e6a57 · inbound
Challenges and Recommendations for LLMs-as-a-Judge in Multilingual Settings and Low-Resource Languages MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 195
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.