Pith. sign in

Paper Citation Record · LEDGER

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

As of 18 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 31 inbound Pith citation observations for arXiv:2505.11942.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11942 v3

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:48:13.467627Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:11:35.363071Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:09:59.374964Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3dbd0a93-d745-4e0f-8d67-78af7a42b733 · outbound

This paper cites GPT-4 Technical Report.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.367578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.367578Z digest=sha256:55dada2646596d9e153c24d206bbec298d8518ec15602ff30b4b32f9d37bc2e6

Observation 3b735d48-060e-421f-afc8-09d01be3c167 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.372321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.372321Z digest=sha256:96570d57a73e1d42bc26d1257073df29f34672671b129dced7dbfa1ee48dcbce

Observation d98d14d9-7e37-4f9a-a53e-6c8a57a54a23 · outbound

This paper cites Minedojo: Building open-ended embodied agents with internet-scale knowledge.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Minedojo: Building open-ended embodied agents with internet-scale knowledge

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.919060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.376982Z digest=sha256:507794d84fb694fbdd21584a9666f1e2544c724f1757094b30be83a89563e98d

Observation a55f3d1a-f802-4692-8250-67a48d1d3a7d · outbound

This paper cites Catastrophic forgetting in connectionist networks.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Catastrophic forgetting in connectionist networks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.380867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.380867Z digest=sha256:9c94df455d27dc27b66edc6fb7e704688c40f0a31c2b6081f8c1e8e3da294c0e

Observation e74e092e-a753-4d4d-8812-6ffd75dd99c1 · outbound

This paper cites The Llama 3 Herd of Models.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.385367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.385367Z digest=sha256:9d9bccbab295d62590a9850f63f5d6969acfc2e1d4a4531652d825d1b598ce94

Observation 50939ffa-f0f5-428d-a6b9-3c587e0b1244 · outbound

This paper cites Sadler, Percy Liang, Xifeng Yan, and Yu Su.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Sadler, Percy Liang, Xifeng Yan, and Yu Su

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.895804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.389812Z digest=sha256:592c26e3f0974e0ea695997bc51daf5d698aeb8789b39bb8932f766984b87f27

Observation 2a942efe-b00c-4544-a47c-89b9b9b53926 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.394573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.394573Z digest=sha256:3f26a93840bd7c1d24af63605ed86ea86a64c453357700954a5aab74191e51e4

Observation aa00b23b-b0c7-4fc8-a23d-31dbf576fa0f · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.398662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.398662Z digest=sha256:f975788139202380c3b4af192e4ae26f6b95fa266e68eed6ddedf8b79bf1f6b6

Observation 73f8705e-ef48-4ff8-a724-8abe1f681f61 · outbound

This paper cites Tree search for language model agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Tree search for language model agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.403617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.403617Z digest=sha256:36ae1872d74f53355870ec5683c032be0a90d678665375045b3787e0b8b89753

Observation fea8fd41-0e54-449d-a32e-0020c3e51d9b · outbound

This paper cites A game theoretic approach to lowering incentives to violate speed limits in Finland.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners A game theoretic approach to lowering incentives to violate speed limits in Finland

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:48:13.669902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.407374Z digest=sha256:22f00b8427d723cf05e28ee342bcbc844ac57d8a245b95cf88790fc0c572edda

Observation 31fd13df-4c97-43d6-9168-548ba699ad90 · outbound

This paper cites From System 1 to System 2: A Survey of Reasoning Large Language Models.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners From System 1 to System 2: A Survey of Reasoning Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.411450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.411450Z digest=sha256:ec81ecac2ca816e268242fe9a1aea200c171936d039f3000d437bf6162d50a7e

Observation 9719d5d0-c10d-45ab-b649-279b965bedde · outbound

This paper cites Cross-Dataset Adaptation for Instrument Classification in Cataract Surgery Videos.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Cross-Dataset Adaptation for Instrument Classification in Cataract Surgery Videos

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:48:13.634197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.415542Z digest=sha256:2b7a555ee23dfe73a2461c6b4bae7ee31abc9066096306c86d1b96bb45d53699

Observation 5a6138c8-0c51-4f60-8392-3dd42cb38eab · outbound

This paper cites VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.419453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.419453Z digest=sha256:936de362799568ffc6bf2537c1385657184a17ffbf3dc5d09edecd17aad69592

Observation 387005ac-794e-41aa-842a-47b827c8b8ab · outbound

This paper cites Large Language Models as Optimizers.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Large Language Models as Optimizers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.423460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.423460Z digest=sha256:26fb71df36e8df92f3a8bd88fc78a818ffc9c5e235596a91a620d63d84f6701d

Observation 47288f4a-5acb-443b-bf48-1ae23320dc44 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Gpqa: A graduate-level google-proof q&a benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.427576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.427576Z digest=sha256:0d2e7cf477e1c26eb8fbf4ca44e3ebc7b6a76d93c21826159113912abda46e8b

Observation fa13e382-22a3-4880-a0f9-a5aaf255f3cd · outbound

This paper cites DebugBench: Evaluating Debugging Capability of Large Language Models.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners DebugBench: Evaluating Debugging Capability of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.431841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.431841Z digest=sha256:cef4fd9207747b52a5997e831d41e3f93894c3d2c1823452be6a867b7ec9811d

Observation d60ec375-a143-4592-84b9-3fe3f9e72670 · outbound

This paper cites Le, Ed H.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Le, Ed H

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.875272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.436152Z digest=sha256:2d2a147d3da0ea80d90e2ac6a605eb4747ddc9f0a277d3e233a47b803547d8f3

Observation f875d5b4-e04a-40e7-9e23-aefd13b3872e · outbound

This paper cites Qwen2.5 Technical Report.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Qwen2.5 Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.441226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.441226Z digest=sha256:5e96842ab851aaba0b06da217f196be746c846963780caf3d8cb00ba5142eb2a

Observation ac6d6d51-f993-425f-b6db-c8320c2bf74e · outbound

This paper cites Agentoccam: A simple yet strong baseline for llm-based web agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Agentoccam: A simple yet strong baseline for llm-based web agents

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.862675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.445450Z digest=sha256:1909dca70867f05ab7c8147db94f0c575ae0acb6cabf0f6f4ef56f1279eccdf8

Observation ae63f2a1-00fa-47c3-bab4-021456f83084 · outbound

This paper cites Towards lifelong learning of large language models: A survey.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Towards lifelong learning of large language models: A survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.449890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.449890Z digest=sha256:6837fec615bd9c3b9ebe79d4cc6c054ec1afbd17be56b2eadc6664fb2da02d2e

Observation 93f12457-9d41-4230-bc38-8e14559d9846 · outbound

This paper cites Lifelong learning of large language model based agents: A roadmap.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Lifelong learning of large language model based agents: A roadmap

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.454126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.454126Z digest=sha256:8cb49416904cc2383f0465c350605a098f730f5f3d2802bc20c9340a5904ae17

Observation 872f8212-6cff-4c8d-ad3b-1e250959e711 · outbound

This paper cites Webarena: A realistic web environment for build- ing autonomous agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Webarena: A realistic web environment for build- ing autonomous agents

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.840248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.458228Z digest=sha256:960aa9586b4118507ceb516ac193a483fbc7274a61c26bc8294f1ec3cdc3eee0

Observation a54097ba-d635-40a4-adf2-3b65dae13f0b · outbound

This paper cites client,".

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners client,"

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.826852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.463287Z digest=sha256:d105a2d8884dd602e319d712f3b0bb136c13c938da62207faa1a74f4cfa1bcc7

Observation 00714672-c52a-44ca-9e31-1dd2281ce3ad · outbound

This paper cites 2023- 10-15.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners 2023- 10-15

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.813203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:48:13.467627Z digest=sha256:5f793db62d2ba665f78d3a57018b38a755faa874139353e0a8873ae1c978c35e

Pith citing papers

Observation 46ed92c7-08f1-4692-b51b-e64cb26ba351 · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 260

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:37.596730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:37.596730Z digest=sha256:83cb8042886e56003fab04da7cacab70ac8f574733d62873050049fb4d50db3f

Observation 934d75bc-9274-4063-a085-9d541639eba9 · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:23:15.776528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:df2c27a5880718c3176294055f904eb13e67aad7d6e4079c64ba1881a415c3a7

Observation 945335f0-3e5d-4dc7-91d9-f2bb909a0158 · inbound

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory cites this paper.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:15.719769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-14T23:13:15.016486Z digest=sha256:8ff2275ab1dcff041c5ee6d39968771b921ed7f3cc1dd6a5bf3176e6bb753476

Observation 102908df-210f-4749-9ad2-ee6c2cb9cf3a · inbound

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems cites this paper.

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:43:12.033070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T22:42:43.070265Z digest=sha256:7a10875bd296e91dc153d4a8cd47f0e46246603cf98e32bac798fc3dc8a8fce3

Observation 40b5de03-7002-4d38-99f1-e954de0aa76d · inbound

LLMs Corrupt Your Documents When You Delegate cites this paper.

LLMs Corrupt Your Documents When You Delegate LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:48:47.713951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T09:47:21.966292Z digest=sha256:947faf78527bc45ac037ed9998548a952c00a3be562cceccd792c3ba9fd55b74

Observation 3bebf6ae-0b6a-4de3-b314-222b4e6c27ac · inbound

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0) cites this paper.

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0) LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:56:47.531182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T06:53:42.607059Z digest=sha256:55549f58cdda1ef796a18b8380f2c10d1ef6ea6d3f2fc9f3307dbf9b39e9a930

Observation 5db456b3-b1ff-4987-8075-ab3858fead24 · inbound

From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work cites this paper.

From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:09.330280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T09:50:30.639962Z digest=sha256:060e85b955fad36dbf13787598ac5108c47d2b683ef798b09923d59afbf1440d

Observation 8ea3b2b9-260a-40f0-a9fb-86bc0edb6419 · inbound

Learning CLI Agents with Structured Action Credit under Selective Observation cites this paper.

Learning CLI Agents with Structured Action Credit under Selective Observation LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:00:56.358550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T02:59:26.100818Z digest=sha256:3075c6ca27ef80ea134b9321203f8ada7bede4568c7598875585771b10ef4f67

Observation a768e694-7ca1-4ee1-9336-ae9873277fe1 · inbound

MINTEval: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems cites this paper.

MINTEval: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:48:12.955233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-20T10:45:08.157521Z digest=sha256:16d0884c474045493fdb3a7ffe76cad55f51a7c66479fdbf2f576b067408e35c

Observation 223a33c4-15ec-4e37-a708-32a2b129ae13 · inbound

Mem-$\pi$: Adaptive Memory through Learning When and What to Generate cites this paper.

Mem-$\pi$: Adaptive Memory through Learning When and What to Generate LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:29:34.495977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T04:27:25.041652Z digest=sha256:50ffd6c36f3a64973e28462e0797e46012adfbc42671ae23197a9ac28f18f674

Observation 7fe3ef37-e2ac-43f3-b5d3-f02eff7d353b · inbound

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation cites this paper.

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:53:47.806029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T17:41:19.270370Z digest=sha256:05ac2b7e3a9b95618782683bfa1a71fc9e5432af8a44b69e76ea3f1fa03bf368

Observation 8e1ee355-1246-4978-bfe0-af7dacd61367 · inbound

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents cites this paper.

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 32

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T22:56:20.140079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T14:52:49.244423Z digest=sha256:3c097ef1ab0e376de3de81f094a80e911af379352023fced02245537ac14f7b4

Observation 68135c1e-bb52-4b5d-a59d-e275a884ae8f · inbound

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents cites this paper.

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.878043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T07:05:18.137509Z digest=sha256:8d5f14c11b6e85864c9c0839b4b1205542528a78ee7e44fd131c400ab7fbbb6b

Observation fde2c070-d2d6-4e65-8ffd-2b2f4c651c5b · inbound

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks cites this paper.

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.802049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T06:16:07.090870Z digest=sha256:cf6fd7366411ac72575a80e4d7ee329dd8e40a40790213cc5458203708b8aa42

Observation fcdb34ad-74df-4b64-a80c-8593f9e9f365 · inbound

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments cites this paper.

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:56:56.846590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T01:45:28.693098Z digest=sha256:1fec04b58737466156262df153840519795f10003b9561dd59510ac1f75e9e7f

Observation e6dea711-ccd8-4f57-93da-c73538e28d79 · inbound

FinEvolveBench: A Benchmark for Self-Evolving Agents on Low-Repetition Tasks with Implicit Rewards cites this paper.

FinEvolveBench: A Benchmark for Self-Evolving Agents on Low-Repetition Tasks with Implicit Rewards LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:09.741298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T22:24:25.502732Z digest=sha256:ac91d69dd6dbf5a6e4e48dab0c0b4d700bdf9df34b54a497e44940a5563ffb8b

Observation 80781c9c-c8a2-4a08-8be4-a2ce2e46fc6f · inbound

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment cites this paper.

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:14.632784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T17:27:19.467192Z digest=sha256:80cf6325d044a918a3aed74f7b19c9e327ba2d0d47ee1f62e1a8ffde200e1148

Observation 99290d27-b412-4094-afb3-2f219213a6d8 · inbound

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses cites this paper.

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:47:25.855088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T19:25:47.085930Z digest=sha256:836967e5ff9e3f6af6393d9ad0583203304fe01ac282b11a3b52c0474854446e

Observation 6178b1cc-eeee-4452-a557-91b63c2ced00 · inbound

From Player to Master: Enhancing Test-Time Learning of LLM Agents via Reinforcement Learning over Memory cites this paper.

From Player to Master: Enhancing Test-Time Learning of LLM Agents via Reinforcement Learning over Memory LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:37:25.575786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T18:48:30.813878Z digest=sha256:ce6a3c8c11d350543ceacada069bf2e2aa3604f06026cf38a602b09ab1b36445

Observation 8fc21509-c1dd-48d0-94ee-0cb1c4e2f20e · inbound

GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents cites this paper.

GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:49:02.507885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T21:45:13.494097Z digest=sha256:70acda71762b5bb0e58af00f218ef52591fb0d610f2ff96f36c7f65566045a6d

Observation d39e33f6-1562-496e-9a34-1abcf0cab0b8 · inbound

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents cites this paper.

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:41.271266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-29T16:55:21.649886Z digest=sha256:4e39c01ede39422cb24f73a30e3848ee224c4c8a6be880638d655262adc5f750

Observation ef7b8b11-cfe1-40bd-837c-96742ad33f77 · inbound

Are We Ready For An Agent-Native Memory System? cites this paper.

Are We Ready For An Agent-Native Memory System? LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:09:59.376525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-25T23:52:26.260258Z digest=sha256:3b1fbc146bb03c6ec1054f3bb60dc834316d7871b7808782b4b6949910bea43f

Observation 76303297-79b6-4d46-bf41-60a5ff39214a · inbound

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents cites this paper.

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:15:48.385777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T03:44:51.320606Z digest=sha256:e80a81c4f741c723247d0d03d5fb8323628b7a12a2acf9e708772b89d014b5f7

Observation a011244a-6d00-4119-a029-e0ee87c7f1a5 · inbound

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments cites this paper.

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-11T07:57:43.000834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:57:43.000834Z digest=sha256:1064e074811f1476bc7baf1c7192119350306375ef4ce3ce7d06bbe7163c4810

Observation 08d995d4-ee66-46e8-abed-1167c6aff41b · inbound

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting cites this paper.

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T11:46:03.282739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:46:03.282739Z digest=sha256:322555c02b3d0c9ec2901cce8db079d9351149ac455c78794e723777b82ffd8b

Observation 884533dd-6334-41ee-af77-a2f36bc6e103 · inbound

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution cites this paper.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.453856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.453856Z digest=sha256:ca8a144e0a1008b468db3b08f618626214f00de7aefbaa02418f3beb99ba89eb

Observation 3f8ec626-822f-4a24-b2a2-620a39c5a29a · inbound

Progressive Multimodal Alignment for Continual Instruction Tuning cites this paper.

Progressive Multimodal Alignment for Continual Instruction Tuning LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-30T16:11:52.369041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T16:11:52.369041Z digest=sha256:e96fc22b2460d864db8b965f0959bd067ff50dc6dca8a733fef05125dc1afc05

Observation 5a7635e6-51f1-4379-9d3d-35c8d1c18e8d · inbound

Progressive Multimodal Alignment for Continual Instruction Tuning cites this paper.

Progressive Multimodal Alignment for Continual Instruction Tuning LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T01:24:03.664375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:24:03.664375Z digest=sha256:ce34d14e8a62004df6637f33e63ef3adc439b1cbb05637916426d5b723a79283

Observation 25d47a20-a4eb-4e00-9e2f-45b721254bb9 · inbound

CoEvo-Mem: Co-Evolving Retrieval Policy and Memory Bank for LLM Agents cites this paper.

CoEvo-Mem: Co-Evolving Retrieval Policy and Memory Bank for LLM Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:11:35.363071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:11:35.363071Z digest=sha256:fc239995249c5497de2171ea440d2c3a38ea7af5a5da5c616771f89cabe60238

Observation 29246f5b-6c86-4555-994b-062debd0b8bb · inbound

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows cites this paper.

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:36.008325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:10:36.008325Z digest=sha256:4c98065177f0fd93a3e43f692e8bb9a98b8f7afafae5e5cdc420fcdbb79839bb

Observation 541eb622-8323-4e1e-b4ea-ccb2a1b59808 · inbound

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World? cites this paper.

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World? LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T15:18:12.475969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:18:12.475969Z digest=sha256:3e141d097751f66225925157515807cb8a9664f12ad7121b391a2ca342b996d7