Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2401.13178.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:32:31.241598Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
8
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f76ff87a-3d72-4686-8967-5ffdc37f8e4c · inbound
OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8fb94b4b-3c84-4b6a-85c1-06bcb76127a7 · inbound
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6202b31-7864-4934-af61-b7d2544c0570 · inbound
Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78642a46-cfb1-4bda-b33f-2afb0c9e06cf · inbound
Make Planning Research Rigorous Again! AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07df1bbd-0184-437d-ac61-5330f2269e25 · inbound
LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc6f3ffc-5f94-4b77-aaed-b64b99c6b270 · inbound
From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 581b98ec-4edd-4610-96ee-2b871f2fa734 · inbound
G-Memory: Tracing Hierarchical Memory for Multi-Agent Systems AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a1d57a7-a4b6-40c4-a890-fc0c4be9b11a · inbound
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 241
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d775f98c-79ce-4fda-ab4b-d2fc7398bc4f · inbound
DrafterBench: Benchmarking Large Language Models for Tasks Automation in Civil Engineering AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59d6be4b-64f8-49cf-977e-956a9ec79f4d · inbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4c10d5-7239-4858-9037-39477c88bd91 · inbound
Evaluation and Benchmarking of LLM Agents: A Survey AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1ba59cb-94fc-4e95-8896-a3cff2265162 · inbound
Galaxy: A Cognition-Centered Framework for Proactive, Privacy-Preserving, and Self-Evolving LLM Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14eba690-db7b-4d04-92b7-fd4fb6d5f73e · inbound
Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 154
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9607fd6-e224-478a-ba99-489caa8123e5 · inbound
Memory in the Age of AI Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 278
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a1aa186-e393-40e7-ac4f-7548b4f78508 · inbound
EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa8b3b2f-1d2b-428a-8fa9-a65bb2bb6c90 · inbound
Sell More, Play Less: Benchmarking LLM Realistic Selling Skill AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6045f6b9-eb30-42c8-a7bb-07f156e290bd · inbound
Holistic Evaluation and Failure Diagnosis of AI Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2aa9f43d-5d04-4293-b1fc-66dad36290ca · inbound
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ddced076-1a45-4ff7-b2d4-29f0fe116742 · inbound
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba622d8f-9203-4871-876f-a11f9308063c · inbound
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1508867e-de20-4be9-b4e3-fbfadd577f49 · inbound
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6de4b676-0d25-434b-804d-9b7c2be9a0eb · inbound
AutoMedBench: Towards Medical AutoResearch with Agentic AI Models AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 702924c7-5ae9-4c68-8ece-9d8b6d499abe · inbound
Do More Agents Help? Controlled and Protocol-Aligned Evaluation of LLM Agent Workflows AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aefd2fa4-d369-4759-95c5-8aa39d835a58 · inbound
Catching One in Five: LLM-as-Judge Blind Spots in Production Multi-Turn Transaction Agents AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 929c8d43-2632-4205-9d91-610cbc9c1482 · inbound
Layer-Isolated Evaluation: Gating the Deterministic Scaffold of a Production LLM Agent with a No-LLM, Regression-Locked Test Harness AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a38e9815-8bdd-43b8-be5b-62b0e1aa79ac · inbound
Do AI-Native Biotechs Need Departments? Benchmarking Company World Models for AI-Driven Drug Development AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 285ed43e-09ab-44c5-bc7e-271c70d3b5ce · inbound
SQBench: A Benchmark for Evaluating Task Delivery by Language-Model Agents in Production-Oriented Workflows AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 078adbd5-018a-4b12-9862-500131aaa417 · inbound
Focus Is All You Need: Adaptive Goal-aware Attention Orchestration for Multi-Agent Graph Systems AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b53841d8-7de1-4d38-8c00-2d2f42968e7d · inbound
WorkSurface-Bench: Benchmarking Enterprise Agents on Multi-Surface Knowledge Routing AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecde07a1-275c-4b25-ba4d-c7856806209c · inbound
Beyond Component Testing: Validating Agentic AI Systems AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9eb8009-1da9-4cfd-b2d0-f72eff0b0898 · inbound
Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.