Pith. sign in

Paper Citation Record · LEDGER

Benchmarking the Spectrum of Agent Capabilities

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2109.06780.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2109.06780 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:18:58.281444Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 82e9da97-b07a-4019-80d4-b67fda08514c · inbound

Mastering Diverse Domains through World Models cites this paper.

Mastering Diverse Domains through World Models Benchmarking the Spectrum of Agent Capabilities

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:08:22.295381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T09:08:21.677362Z digest=sha256:0973e430205757159ece2109af64b41fd1d5052c12dc2afa99a3832467d012ff

Observation 3a487a79-7bc9-4b6b-9649-e5130d2c9448 · inbound

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution cites this paper.

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution Benchmarking the Spectrum of Agent Capabilities

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:12:31.551279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-16T08:12:30.984870Z digest=sha256:fecd29a4b9a92582aa4dd960eb8859772a8c9b2c65ef83a6a2624eea7e859490

Observation 3a91a203-1efd-4d38-918f-ef197a4e667c · inbound

Disentangling Exploration of Large Language Models by Optimal Exploitation cites this paper.

Disentangling Exploration of Large Language Models by Optimal Exploitation Benchmarking the Spectrum of Agent Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T20:18:58.281444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:18:58.281444Z digest=sha256:e4352579b8e19f8e01d7574926831c3dcbfc1c3d3a650a6a04e24906c11c7fd8

Observation 1b4d7779-3cc1-4ea8-9941-d73cb98bcc66 · inbound

Improving Transformer World Models for Data-Efficient RL cites this paper.

Improving Transformer World Models for Data-Efficient RL Benchmarking the Spectrum of Agent Capabilities

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T14:59:44.910426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:59:44.910426Z digest=sha256:f39a312153234eb2bf291a61097dbe6ed46893b1360022149e8db6c165a925c2

Observation b75dd297-58d9-46a3-8584-5bba21f5289a · inbound

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games cites this paper.

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games Benchmarking the Spectrum of Agent Capabilities

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T12:02:16.686722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-19T12:01:42.681135Z digest=sha256:ac2f0df032a21e098ad597c20e86eece4f8a95b83d83f3800ce1ba9c94852376

Observation f27307be-998b-4f03-aa03-d7157a839832 · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey Benchmarking the Spectrum of Agent Capabilities

Reference 181

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.829118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.829118Z digest=sha256:f4dc0e8220793cd21fa1b026ad8d6d012e5bd9d9c419a0d9c4b5924bef70fd0d

Observation 46d342f4-049b-438c-9cb7-c743d4c9a386 · inbound

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley cites this paper.

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley Benchmarking the Spectrum of Agent Capabilities

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:21.183895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:21.183895Z digest=sha256:7bdad7424226f83767ca3b5df5400bb637370e3b5b271158e4119f028a1e31d6

Observation 0d14a9d1-f1ad-4f76-9b59-997042574c7c · inbound

How Should We Meta-Learn Reinforcement Learning Algorithms? cites this paper.

How Should We Meta-Learn Reinforcement Learning Algorithms? Benchmarking the Spectrum of Agent Capabilities

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T14:48:47.078571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:48:47.078571Z digest=sha256:03ddc9014fc361fbbd160d7d97f2f66935ff95f7e1ba5b5aa03c014a739e203e

Observation 434a26f0-1083-433d-bfd8-94779226a9e4 · inbound

Agent-centric learning: from external reward maximization to internal knowledge curation cites this paper.

Agent-centric learning: from external reward maximization to internal knowledge curation Benchmarking the Spectrum of Agent Capabilities

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T11:58:03.781339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:58:03.781339Z digest=sha256:2376bee044e9b1809483d6ed0405e5f275e4914ca21e53b2e2d3a3c65ff9aa0f

Observation ca241428-bd87-4696-8b30-f0759e9f77ce · inbound

Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains cites this paper.

Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains Benchmarking the Spectrum of Agent Capabilities

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T10:34:35.501264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:34:35.501264Z digest=sha256:7d2c03b569e9eff8b070df6475cafc0b3de01aa33b02bf635c42c51cdbf99a8a

Observation afd4d3b6-33c3-4d96-8c29-c5fe411cd001 · inbound

Accelerating Reinforcement Learning Algorithms Convergence using Pre-trained Large Language Models as Tutors With Advice Reusing cites this paper.

Accelerating Reinforcement Learning Algorithms Convergence using Pre-trained Large Language Models as Tutors With Advice Reusing Benchmarking the Spectrum of Agent Capabilities

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-04T20:49:37.895334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:49:37.895334Z digest=sha256:7c3a3a8982770fc663b913f1d9ba78e7efcce4d2c99052ba8450035486276a4e

Observation 5939e947-2c33-46a4-bbd6-9b43e4daf45b · inbound

Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX cites this paper.

Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX Benchmarking the Spectrum of Agent Capabilities

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T12:54:40.079520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:54:40.079520Z digest=sha256:d43c7dea83792e2d145c027b788b0c1e4c8b7b467d34b2cf1d48ac04de40df8c

Observation 45788988-8799-41b1-a21a-f24b883bc20c · inbound

BuilderBench: The Building Blocks of Intelligent Agents cites this paper.

BuilderBench: The Building Blocks of Intelligent Agents Benchmarking the Spectrum of Agent Capabilities

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-04T11:20:18.762379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:20:18.762379Z digest=sha256:bb3640edb7c934fba0b5dd198fd119ac5a9a4d87450e7369829c6b170bb7ecb4

Observation 7f293be8-eca4-43f2-8267-4abeb31ed2a3 · inbound

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction cites this paper.

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction Benchmarking the Spectrum of Agent Capabilities

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:50:05.072031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T14:46:18.645009Z digest=sha256:a595c57337104c080477460670fee7650e4b52c9a6ad96b4d0a67e3199ac594f

Observation 830b4ef6-7057-48b6-b4c7-d62b5b32b340 · inbound

PACE: Parameter Change for Unsupervised Environment Design cites this paper.

PACE: Parameter Change for Unsupervised Environment Design Benchmarking the Spectrum of Agent Capabilities

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:01:05.807753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-09T14:21:34.002785Z digest=sha256:c0c65cdfbd8a267aeb881c36bcc1b7cabe0883975c1d2a63cf9a759bb380a286

Observation e879502c-7fcc-4936-aac0-9abf3e28be70 · inbound

TRAP: Tail-aware Ranking Attack for World-Model Planning cites this paper.

TRAP: Tail-aware Ranking Attack for World-Model Planning Benchmarking the Spectrum of Agent Capabilities

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:04.939841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T15:15:45.407680Z digest=sha256:123c13335ee17675d571401df9a38250afb91fa2d34b345cb259d68b231897d3

Observation 9ae4083e-ac65-43ac-ae04-abd6a9a62ad0 · inbound

MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs cites this paper.

MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs Benchmarking the Spectrum of Agent Capabilities

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:26:25.687344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T05:13:28.089038Z digest=sha256:2f12c99057e1b4693d8385af64e8d4b178f909a1c113ae9fa6627f936d1d3704

Observation bfa3be0f-3892-42c5-836d-3e4ae9cced7d · inbound

CA2: Code-Aware Agent for Automated Game Testing cites this paper.

CA2: Code-Aware Agent for Automated Game Testing Benchmarking the Spectrum of Agent Capabilities

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:55:05.012292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T05:53:26.558042Z digest=sha256:3f5f02ba5685b3ef26e38e66b2c71989126403cec07352f79a499da430f42037

Observation 84350e10-0c1e-4d68-8618-95122ff2fc3d · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders Benchmarking the Spectrum of Agent Capabilities

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:33:04.172830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T05:28:50.354662Z digest=sha256:3a4a9d2f0eef8d2b03944fb7170940c7127ce48f844126cdebaf9821172fb539

Observation c2164140-34cc-4f18-9e2e-7da0d23eb1e3 · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders Benchmarking the Spectrum of Agent Capabilities

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:39:49.261351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T07:36:12.214949Z digest=sha256:4f5a476f0410b6317c31fb67cd303dfb0f3b64836a553bb94cbb43bdabfe999e

Observation 80a7b6c3-2c47-4eb9-866c-864ba20e75fa · inbound

Goal-Conditioned Agents that Learn Everything All at Once cites this paper.

Goal-Conditioned Agents that Learn Everything All at Once Benchmarking the Spectrum of Agent Capabilities

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.614820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-25T04:59:48.867927Z digest=sha256:2a5088f76413d3c676b429fe53d33a6e7ad7aaf029936c4090faf5c34c7c0681

Observation 6861cde4-1217-47d5-80a1-c0ce1ab3b807 · inbound

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness cites this paper.

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness Benchmarking the Spectrum of Agent Capabilities

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:33:28.520711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T13:23:32.865797Z digest=sha256:f743337fd324136139024ba9636211eadb2a5466e2a76abf0130a1f1b3b02b27

Observation 2561077d-70e5-4abf-9ac2-d47207b48484 · inbound

CLAW: Learning Continuous Latent Action World Models via Adversarial Latent Regularization cites this paper.

CLAW: Learning Continuous Latent Action World Models via Adversarial Latent Regularization Benchmarking the Spectrum of Agent Capabilities

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:36:29.557041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T09:55:00.402411Z digest=sha256:82cbb1469c2a9629cbab695cacc2874865a3db112030a05c796bcd3442685416

Observation e6aa97de-3d72-4584-a615-0d315db61e88 · inbound

APPO: Agentic Procedural Policy Optimization cites this paper.

APPO: Agentic Procedural Policy Optimization Benchmarking the Spectrum of Agent Capabilities

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T09:37:49.433192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T10:21:55.485624Z digest=sha256:16f39d2d6f69996f9aeb4f0b084e8bc301785f2532afb7aaff5192ba4c6bbf93

Observation 1a9f0cd7-186a-4cc2-a412-2435c0e2451c · inbound

APPO: Agentic Procedural Policy Optimization cites this paper.

APPO: Agentic Procedural Policy Optimization Benchmarking the Spectrum of Agent Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T02:12:26.921207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:12:26.921207Z digest=sha256:a10e091f319f33cbba3600136725cec48f98fc77a8f5bd7584f9cfe049740f23

Observation fd34091b-73b4-464b-9498-b5329877e36b · inbound

AutoMem: Automated Learning of Memory as a Cognitive Skill cites this paper.

AutoMem: Automated Learning of Memory as a Cognitive Skill Benchmarking the Spectrum of Agent Capabilities

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:16:56.527740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-02T12:09:32.053764Z digest=sha256:cb07592a3bdeb062a11b5b9c24ceb68d8fac93834dab5a2faf97d3f938f399df

Observation c762bc2e-bd9d-408f-a772-383437155fa3 · inbound

DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat cites this paper.

DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat Benchmarking the Spectrum of Agent Capabilities

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T04:15:05.687146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:15:05.687146Z digest=sha256:ad36e9a4ad38f803c8b2fdbd62ba32dd430255cc7d06db728f06093de1d3fe1e