Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:46:02.635196Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 2 inbound Pith citation observations for arXiv:2507.12806.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:46:02.635196Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-21T15:10:04.253250Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T15:10:16.398252Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 965e8d88-5920-4fa0-997d-18671108562e · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f48c8dec-5f67-4714-8de1-2e2eb14d62c1 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54a68f72-8b63-4ba7-9b90-4afa6c213840 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0ac6f1bd-aad2-4886-b45b-2511928efe4b · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f2ccccd-f0fe-44b7-975a-a4917b8daabc · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models On extensions of the Jacobson-Morozov theorem to even characteristic
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ddcdea5-a3c5-4978-b407-ccd568d67cb5 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff70678c-e4b2-44c1-80ea-76173553b1ad · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 382aea7f-4f3e-4d31-8b83-8ccda811beb7 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models High-temperature oxidation and nitridation of substoichiometric zirconium carbide in isothermal air
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9963f502-26c4-4230-be21-15f6fc5160b3 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a58901e-2fd1-4e76-b9f1-3f512b7b46ab · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models REALM-Bench: A Benchmark for Evaluating Multi-Agent Systems on Real-world, Dynamic Planning and Scheduling Tasks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52fde6de-9293-4f89-9276-73de583e86ea · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Measuring Massive Multitask Language Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a05bf665-2307-4f98-84fe-e0c110d1fff4 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b413d103-a494-4f61-978e-3a676697e357 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Recurrent neural chemical reaction networks that approximate arbitrary dynamics
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c2a0c8e-840a-49e8-a759-c6788b9e8af9 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5e270d4-5040-4774-9917-b0b1f1891b6b · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da1ec4f5-2cea-4d02-9714-852c90562a85 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e3145b2-54c9-4e9e-b271-36a5d95ca7b4 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f16da7f4-79b5-4715-9121-e8d03a7cfd06 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39716159-e2c8-490a-8800-1e82ea79b651 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models AgentBench: Evaluating LLMs as Agents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10d30c61-c805-483a-aca3-821be6d5c8e2 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models PRACT: Optimizing Principled Reasoning and Acting of LLM Agent
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1500648-c6dc-4fb5-bec1-bc9d37bc47dd · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b85cbb9-9ee2-4265-8954-2dbfe0457389 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3533abeb-e1d2-4a2d-a2b4-fe04b0e9cc98 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5bacc67b-035f-41f7-9ede-cda04be96e11 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models ScaleMCP: Dynamic and Auto-Synchronizing Model Context Protocol Tools for LLM Agents
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59d6be4b-64f8-49cf-977e-956a9ec79f4d · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 601a5aac-4ffb-4790-836f-c4219bfdb3fc · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b9de82d-5851-4695-a7c3-af6f7b0f7fc1 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models GPT-4 Technical Report
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f61e1be-832c-47c5-ab38-ffe3057f52e3 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation afb76a43-b3a7-42b7-aa5e-5c8f1f69623b · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a46a96-655c-4937-a520-f84420cf7735 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 277c3b40-f282-435e-8481-dc53cbe330fc · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cea0ddfc-24f6-4de2-b2ae-41327a4f6009 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models PersonaBench: Evaluating AI Models on Understanding Personal Information through Accessing (Synthetic) Private User Data
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c31e5581-2880-462d-af0b-097ac6237e88 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Peak Age of Information under Tandem of Queues
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91de0b95-21ad-4e1d-ad2b-f137266aff4e · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f13d3aeb-25cb-4548-b755-eafdb31c03dd · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fe92b4f-1fe3-41e0-8586-4fe378d8f71b · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models LiveBench: A Challenging, Contamination-Limited LLM Benchmark
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b196f587-3af0-4831-97d9-a95dceff3733 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6aaf830b-3ec9-4809-b1b8-55ab9db3030f · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24645806-899e-471a-9159-5064ce9a8a91 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models WizardLM: Empowering large pre-trained language models to follow complex instructions
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1eb775e8-6a66-43bf-8819-92406bbd28fe · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models MCPWorld: A Unified Benchmarking Testbed for API, GUI, and Hybrid Computer Use Agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cdc92fc-d07f-4d04-9a2a-05d6c3adb33f · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a3b0ae8a-5fc8-465d-bbed-990c8b4dfcc3 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models A Survey on Large Language Model based Autonomous Agents
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a242dc43-64ca-404a-a3e3-fadb42c1a74e · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c607c54f-dc16-4cb5-bf78-d021f4fff877 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models ActionStudio: A Lightweight Framework for Data and Training of Large Action Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 380140d5-000a-4b40-ba1e-8f10e9c4d317 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models xLAM: A Family of Large Action Models to Empower AI Agent Systems
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbf4f137-366a-4b92-904e-d90b7f7e9806 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models DialogStudio: Towards Richest and Most Diverse Unified Dataset Collection for Conversational AI
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 247dbebb-307f-41a6-a3ad-76cf635dbe64 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Laser Printing of Silver and Silver Oxide
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bdbe1501-a723-46ee-a294-bc95b9e765c1 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccbb6752-8574-4506-9e62-d15cdad40e39 · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51c148d-fa70-4e16-babb-41243255782c · outbound
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Vec-Tok-VC+: Residual-enhanced Robust Zero-shot Voice Conversion with Progressive Constraints in a Dual-mode Training Strategy
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f50b43b1-2286-4179-9d51-536a3c89175b · inbound
MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 639cab6f-2318-46d1-b552-f2ad2c5d1ae4 · inbound
MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.