Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:51:19.859577Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 6 inbound Pith citation observations for arXiv:2506.17335.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:51:19.859577Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T12:32:39.706055Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:49:57.838826Z
12 of 12 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 57159559-538c-43d7-8ad9-a324dc09a8fe · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 87d77231-adf4-41bb-96d3-b30f3d6139cd · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f33282d9-4110-4476-a0ca-49dfe6ab5815 · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 883b6106-d5c8-4d33-95d9-44f1c38e3cb9 · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research ML-Bench: Evaluating Large Language Models and Agents for Machine Learning Tasks on Repository-Level Code
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9a25c14-4a9a-4ea1-a8db-50e3ec9c6a9b · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research PaperBench: Evaluating AI's Ability to Replicate AI Research
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af802a7a-dda6-45ba-8d27-98008dd0d2fc · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation baee7219-cc49-4df9-ba1a-11a586e082d7 · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 09dd6fd4-f00d-4c73-939f-63316f0f3191 · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8b116405-54ef-45cd-bdcf-025233dcdeb7 · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc2862dc-3e49-4de9-b15c-44655210cd94 · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Evaluating Large Language Models Trained on Code
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14d38a06-45a6-4507-820f-6cbd5426ef1b · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Qwen2.5-Coder Technical Report
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be67592c-3bc9-444c-85d6-8b9f3e0caeba · outbound
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research Scholar Inbox: Personalized Paper Recommendations for Scientists
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e6d145cd-0a8e-4942-83c0-02e112f8beab · inbound
ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 60fbff89-35b2-4fa3-bcb4-5cb540e0de55 · inbound
SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences? LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3ced7e92-6834-4b28-94c4-b79fb9cdeeea · inbound
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0d1807e1-3a13-4bf4-93c1-7555d34ff371 · inbound
AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d85ef8c4-0c1f-483c-b4cc-b16d2bd675b5 · inbound
NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers? LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c805db18-cecd-469b-8f1a-542ec7d94869 · inbound
NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers? LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.