Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T18:53:05.214761Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 4 of 4 outbound references and 0 inbound Pith citation observations for arXiv:2508.13938.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T18:53:05.214761Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
4 of 4 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7464e93f-84fc-44e9-a514-0b3aa0c65612 · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac95634-5e48-4fae-897d-b95e50e70bcc · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models Evaluating the Performance of Large Language Models on GAOKAO Benchmark
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f8d49f-094c-4fea-85e1-5baa35d2c0d5 · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models A Survey on LLM-as-a-Judge
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2498c15-1ae6-4456-91ac-4b9828478eed · outbound
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models R-Bench: Graduate-level Multi-disciplinary Benchmarks for LLM & MLLM Complex Reasoning Evaluation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.