Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T18:53:05.214761Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 4 of 4 outbound references and 1 inbound Pith citation observation for arXiv:2508.13938.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T18:53:05.214761Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:12:45.523967Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-10T18:12:46.678251Z
4 of 4 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7464e93f-84fc-44e9-a514-0b3aa0c65612 · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac95634-5e48-4fae-897d-b95e50e70bcc · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models Evaluating the Performance of Large Language Models on GAOKAO Benchmark
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f8d49f-094c-4fea-85e1-5baa35d2c0d5 · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models A Survey on LLM-as-a-Judge
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2498c15-1ae6-4456-91ac-4b9828478eed · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models R-Bench: Graduate-level Multi-disciplinary Benchmarks for LLM & MLLM Complex Reasoning Evaluation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e4767a2-f4f0-40e0-856d-c10dc52e55f6 · inbound
Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.