Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2402.12348.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T19:37:15.933653Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 45209f9c-2cd0-447c-90b2-50cebb81d5f7 · inbound
Estimating LLM Uncertainty with Evidence GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d822c53b-83bc-4bc7-897f-ac9d279a87f7 · inbound
Verbalized Bayesian Persuasion GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 117a72f0-ac6a-4a25-9b3a-786bc51021ee · inbound
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccdb28e0-7a0d-4687-90bb-27dd9651a7fe · inbound
VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 10149eab-6e4a-41bf-af76-6a916325d0b8 · inbound
Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d37a19b2-dfde-4166-84ee-8fc49b616ab4 · inbound
WGSR-Bench: Wargame-based Game-theoretic Strategic Reasoning Benchmark for Large Language Models GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ecc2b31-fb22-4aec-984e-ffa1694e5e8a · inbound
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 114
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2794b4f1-eeb1-45e2-bcbb-2ac90dc8244e · inbound
Tracing LLM Reasoning Processes with Strategic Games: A Framework for Planning, Revision, and Resource-Constrained Decision Making GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88deb9cc-21ec-4eea-a3d7-eeea63417592 · inbound
The Effect of State Representation on LLM Agent Behavior in Dynamic Routing Games GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ade39c0f-f51e-4c82-8f16-222cdfb2ae79 · inbound
DipSVD: Dual-importance Protected SVD for Efficient LLM Compression GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 025f81de-cd78-4010-a445-bc1cccddaa54 · inbound
How Large Language Models play humans in online conversations: a simulated study of the 2016 US politics on Reddit GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 611fddd2-c4d0-4182-8f8c-7c17f856217b · inbound
GenAI-based Multi-Agent Reinforcement Learning towards Distributed Agent Intelligence: A Generative-RL Agent Perspective GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30f59d22-4284-4e1c-a3ca-b9322be083f2 · inbound
Democratizing Diplomacy: A Harness for Evaluating Any Large Language Model on Full-Press Diplomacy GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12bcbc80-d732-47ad-be20-fc098b3ac09b · inbound
Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9f2990fc-a0f6-4512-8b83-4b798bd08e2d · inbound
Common-agency Games for Multi-Objective Test-Time Alignment GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2f7e3d0e-74a3-4145-8dd5-c89289876e4b · inbound
DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aa4058ae-4165-46e4-a750-1b83aee8f181 · inbound
Robots Need More than VLA and World Models GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 143
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d50edb7b-404d-45ca-b0c9-ac88b352d22b · inbound
Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c59a98e-472a-4f27-a2aa-27ca3b39248e · inbound
Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b292097a-1f13-4384-a22d-90153f7ca041 · inbound
BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 45aeb99a-2a1f-4e2d-8263-e1c954e94bb1 · inbound
Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 71dc7915-1b89-4660-a538-38354c769283 · inbound
Towards Reliable and Robust LLM Planning: Symbolic Feedback-Driven Iterative Self-Refinement Framework GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d655b8b7-d299-4544-a174-9373c020718b · inbound
Scaffolding the Strategist: Architecture-Dependent Reasoning Interventions in Hotelling Spatial Markets GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ee0253-bde2-48e4-acba-bb511d2da326 · inbound
Strategy, Not Payoffs: A Behavioural Embedding of Normal-Form Games GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.