Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2503.01935.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:56:36.558074Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T03:27:35.774566Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a61c6093-c986-49db-9947-37c7a3aeb1ab · inbound
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 27510b94-8cf9-4669-8915-68833004d2e4 · inbound
Know the Ropes: A Heuristic Strategy for LLM-based Multi-Agent System Design MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f9e543e-bfbe-4e9e-87ff-8f1bac9195ca · inbound
MAEBE: Multi-Agent Emergent Behavior Framework MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0b45b05-2cb2-4c42-a962-c81f9586c150 · inbound
$\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c3d80f2-62f4-4044-bb31-68a94500eb03 · inbound
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 153
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b7069b3-c9b9-4bfe-aff9-bdfb63238d12 · inbound
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 161
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d02d8bf0-e065-4fe5-8a9a-ad4782b7ba78 · inbound
Agent Identity Evals: Measuring Agentic Identity MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c99c668f-8563-40b6-b67b-6700452afdfd · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8170dce8-26b6-470c-ac13-5f6f89c498e4 · inbound
UserBench: An Interactive Gym Environment for User-Centric Agents MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfef6379-2f38-4db2-b083-513148057539 · inbound
Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8d35b3e5-4ab4-4ab1-a102-a5fa0cdcab1c · inbound
Agentic Reasoning for Large Language Models MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 18a3958e-d235-45ee-9b01-d126a6916f3e · inbound
Choose Your Agent: Tradeoffs in Adopting AI Advisors, Coaches, and Delegates in Multi-Party Negotiation MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d15010e0-8901-4f6b-8194-cf257e3f1519 · inbound
OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2881e264-9ac3-46c6-abf5-762bfc09a607 · inbound
SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbcb25a6-0a15-4e04-9a92-5f7ab06450ae · inbound
SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0955dcc-c49e-4551-abb5-d0876c56fec9 · inbound
AgentSocialBench: Evaluating Privacy Risks in Human-Centered Agentic Social Networks MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 252e6b62-495a-41a3-9e60-04fd99f068d2 · inbound
AI scientists produce results without reasoning scientifically MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bfe52e0a-8231-4ace-9ecf-4359059f2713 · inbound
TeamBench: Evaluating Agent Coordination under Enforced Role Separation MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7372c09c-75cc-441f-be0a-3fb26c719e42 · inbound
TraceFix: Repairing Agent Coordination Protocols with TLA+ Counterexamples MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 00aa1280-cc3c-445d-840c-a5a1a8a3157b · inbound
Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 10c0f2f7-b6ee-4420-803c-3172a489afa9 · inbound
AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7773379c-34cc-48ac-8804-23ed0c0d0fca · inbound
FASE: Fast Adaptive Semantic Entropy for Code Quality MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50ca8191-5879-40c2-9331-230a41baadee · inbound
Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a7326bcc-83b5-4437-9451-6f157cfe7998 · inbound
Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4fceb272-b46e-48c0-829e-7c6da2ea20d3 · inbound
MAS-Lab: A Specification-Driven Validation Framework for Reliable Multi-Agent Systems MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a261d788-53f4-44ef-8fe5-4c0af4caab2d · inbound
AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff51b4c0-9728-4b8b-a89c-5779826742ec · inbound
The Cost of Knowing: A Resource-Aware Protocol for Benchmarking Hallucination Beyond Static Leaderboards MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9d2bbf0-6bf3-475e-99a6-2a48752b9d5b · inbound
Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.