Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2508.15361.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T12:32:11.218878Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T05:09:36.772485Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 07fe1144-ddf9-4df4-b08b-c07c736bcc2d · inbound
AfriEconQA: A Benchmark for Quantitative and Temporal Reasoning over World Bank Economic Reports A Survey on Large Language Model Benchmarks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f724800-82da-4ebd-abcf-5d6369148920 · inbound
Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization A Survey on Large Language Model Benchmarks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db88210b-8749-4036-9c9c-4d316f149e1e · inbound
The Necessity of a Unified Framework for LLM-Based Agent Evaluation A Survey on Large Language Model Benchmarks
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be0d1a4e-6ed5-4262-b13f-5a01d2087c8a · inbound
When control meets large language models: From words to dynamics A Survey on Large Language Model Benchmarks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21e40198-d230-4747-be18-015923311a6e · inbound
Designing for Error Recovery in Human-Robot Interaction A Survey on Large Language Model Benchmarks
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c65da8d6-9ea8-4c10-819d-5b3c4d5b5b60 · inbound
Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints A Survey on Large Language Model Benchmarks
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 348809dc-1ec0-4909-ab9c-5f398ad4fd73 · inbound
Efficient Ensemble Selection from Binary and Pairwise Feedback A Survey on Large Language Model Benchmarks
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 902288b6-666d-4eca-9fbe-18693aa5ac5a · inbound
Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels A Survey on Large Language Model Benchmarks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12885f74-9760-4eeb-b52d-540af269d061 · inbound
LPDS: Evaluating LLM Robustness Through Logic-Preserving Difficulty Scaling A Survey on Large Language Model Benchmarks
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a896fbe4-8c19-4610-b21a-9f6d2fe46397 · inbound
How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness A Survey on Large Language Model Benchmarks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca851cbb-deb2-42f9-b633-41d537c38d3c · inbound
Can LLM Teams Play What? Where? When? A Survey on Large Language Model Benchmarks
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a497ae1-a56e-4b3c-a908-df4ec41bc80c · inbound
MAVEN: Improving Generalization in Agentic Tool Calling A Survey on Large Language Model Benchmarks
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b4546810-6750-4ab7-ab34-ec27b1103b63 · inbound
Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting A Survey on Large Language Model Benchmarks
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2dab6c41-64f5-493a-9cd3-c7222a0710bf · inbound
Token-Operations-Oriented Inference Optimization Techniques for Large Models A Survey on Large Language Model Benchmarks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 346564ca-59ff-4047-9b06-31b8fc02ed89 · inbound
Token-Operations-Oriented Inference Optimization Techniques for Large Models A Survey on Large Language Model Benchmarks
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5ae7f1e-aa0e-4b00-962f-d06f9c6a82b3 · inbound
YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese A Survey on Large Language Model Benchmarks
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.