Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2410.09893.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:36:45.894072Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-25T06:06:43.201973Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d474fa82-b0e3-4186-8838-bca08a1cde27 · inbound
Qwen2.5 Technical Report RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44a8b972-e68a-4d35-9937-36f6c758c184 · inbound
RewardBench 2: Advancing Reward Model Evaluation RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be8ff1ed-8400-433b-aa1e-e13b8a79b760 · inbound
Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87686c92-4e7f-4a09-a47c-0fcb6213e753 · inbound
CompassJudger-2: Towards Generalist Judge Model via Verifiable Rewards RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d4aa580-82cb-4bd6-ab26-895ceb6edbdb · inbound
URPO: A Unified Reward & Policy Optimization Framework for Large Language Models RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59d294e0-8a4e-4220-8669-a01adfbffc31 · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 126
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 637b1400-4ceb-4add-9383-c99153db5a94 · inbound
AI Can Learn Scientific Taste RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 552759f0-cecb-4c3b-b700-587d580dc87b · inbound
Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 70897c33-a790-4525-8ca9-29f8cadad99f · inbound
Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5900649-fd9b-4461-a55e-64c5f27f6854 · inbound
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 69088732-1197-438e-a9d6-2f62f6f64f58 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9b7a6ff5-e858-41a9-b9f5-3931a8b0bc39 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety RMB: Comprehensively Benchmarking Reward Models in LLM Alignment
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.