Pith. sign in

Paper Citation Record · LEDGER

Scaling Test-time Compute for LLM Agents

As of 10 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 16 inbound Pith citation observations for arXiv:2506.12928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12928 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:41:48.421514Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:20:42.130163Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T06:24:40.465949Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ead75c48-fe3f-489f-b8b7-8104120d74c8 · outbound

This paper cites an unresolved cited work.

Scaling Test-time Compute for LLM Agents Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:41:49.377187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:41:48.279342Z digest=sha256:f7d15d2d616455b0f66a2f092de708ef5e17e63ea9a5ea72aa8769ce9c7ee23c

Observation f453d572-e8d7-42f0-b855-03bab1520a0b · outbound

This paper cites TapeAgents: a Holistic Framework for Agent Development and Optimization.

Scaling Test-time Compute for LLM Agents TapeAgents: a Holistic Framework for Agent Development and Optimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.283939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.283939Z digest=sha256:9817438a17c0e39db875163908bad359670397f2cdf6626243b4e120c5459a92

Observation 16d899cb-5f15-49f7-84dd-93289ab9351c · outbound

This paper cites Beeching, L.

Scaling Test-time Compute for LLM Agents Beeching, L

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:49.360522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:41:48.288676Z digest=sha256:d287bd307250121778aba102fcd8f38d486573486bd52cf38543f5aa97a4e14e

Observation 5e5c012c-e752-40d3-b687-f9e5a26ffd5c · outbound

This paper cites an unresolved cited work.

Scaling Test-time Compute for LLM Agents Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.293050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.293050Z digest=sha256:0be9724200435979fb6a30157603d9b12523c3648eafd4b12d0bdaecbf47b42a

Observation f08d72e9-e786-4ef7-ae03-19a4b8ca0356 · outbound

This paper cites an unresolved cited work.

Scaling Test-time Compute for LLM Agents Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:41:49.345844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:41:48.297183Z digest=sha256:d11587d80bcf9cefda99e97ea0e65cf3d4fc1e87d33acd62d5937a05c0a5c642

Observation 959bf115-9cbd-424c-b39b-23232d1d270a · outbound

This paper cites Faria and N.

Scaling Test-time Compute for LLM Agents Faria and N

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.301350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.301350Z digest=sha256:b84b5e87f9dbf3fa794e7e137d0d8dbdd1825b20c79345fbae284610076e6bdf

Observation 9e291ef7-b138-4076-91b6-2f0d96ab0c20 · outbound

This paper cites Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks.

Scaling Test-time Compute for LLM Agents Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.306102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.306102Z digest=sha256:bd4cf7b3f699c551e6436809c33aa54b886be08aed64d145f308bcc06077bcf1

Observation c2331d8e-077d-446b-98f9-f65c1da3dc75 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Scaling Test-time Compute for LLM Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.310739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.310739Z digest=sha256:03276520945480bea6d0e849a1f37e602cd648bf4162eb09f39cacc551aae8ed

Observation 1684e9e7-5946-4bd6-b138-2cc313d9364c · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

Scaling Test-time Compute for LLM Agents MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.319126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.319126Z digest=sha256:2958d6f2504127383dfb8ce780daee314f4ea0c038fe1f4ea8ec37393c87057e

Observation 3ca74235-d0d8-4eae-807c-f11b019e2462 · outbound

This paper cites OpenAI o1 System Card.

Scaling Test-time Compute for LLM Agents OpenAI o1 System Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.323210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.323210Z digest=sha256:d51c408675bb9921b76536b895d9d879068876af6c420710ca7bbcba17597730

Observation c63f3175-2472-4843-9ce2-89045992f51e · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

Scaling Test-time Compute for LLM Agents VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.327648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.327648Z digest=sha256:28d8b1a0cafeb514ae2e21d48be74fdacc974deb14cb2d4e329a40299883e8c0

Observation 1288602c-0046-48da-bdbc-eaa49d75aa3a · outbound

This paper cites an unresolved cited work.

Scaling Test-time Compute for LLM Agents Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.331771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.331771Z digest=sha256:4a7fd46372e6a41197e51db1764c971c2b97a771899e32b061cb22bd01a8f068

Observation 439ff490-1b63-40a1-866d-b56c2015fab4 · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

Scaling Test-time Compute for LLM Agents Training Language Models to Self-Correct via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.335591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.335591Z digest=sha256:d2c5bfef296868824264d90a6f7630a0ad44987ffab77a086dfb46854e72b504

Observation 21827c23-70e0-4de2-9183-1a3345be1115 · outbound

This paper cites an unresolved cited work.

Scaling Test-time Compute for LLM Agents Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:41:49.330821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:41:48.339706Z digest=sha256:78b43449538522b92f270e7708a7765c1023ccfc9edd8c67461248bff382a197

Observation ebd40cb0-8f36-42fb-889f-eaa512995a70 · outbound

This paper cites Search-o1: Agentic Search-Enhanced Large Reasoning Models.

Scaling Test-time Compute for LLM Agents Search-o1: Agentic Search-Enhanced Large Reasoning Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.343427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.343427Z digest=sha256:f6dba3c5ce58be25954bb5aafea2828de6756853d8acf8e7957f3abb5236b7cc

Observation 6d691e17-8845-48d9-9b29-8ad550720560 · outbound

This paper cites WebThinker: Empowering Large Reasoning Models with Deep Research Capability.

Scaling Test-time Compute for LLM Agents WebThinker: Empowering Large Reasoning Models with Deep Research Capability

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.347863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.347863Z digest=sha256:529bcc5b3d871d740203653d7687526b72ed8e08769c2874a3534f13540e3f45

Observation fb48ef7b-99d4-4531-ab86-8dda35c847a0 · outbound

This paper cites Liang, J.

Scaling Test-time Compute for LLM Agents Liang, J

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.352324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.352324Z digest=sha256:c77e426b95317d3d18fae4f0bb5e48c76b0130b6d0b29eb6e948d93c80a13690

Observation e893122c-ad37-4a27-a137-ea9315d31a5f · outbound

This paper cites Bag of Tricks for Inference-time Computation of LLM Reasoning.

Scaling Test-time Compute for LLM Agents Bag of Tricks for Inference-time Computation of LLM Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.356267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.356267Z digest=sha256:2a268d7cd3218c4c12606dcc9c6dd24c66c22a2412614e0f510f5b7677568fba

Observation b527d30a-272e-4fb3-9d5c-64ecec4f111a · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Scaling Test-time Compute for LLM Agents Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.364617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.364617Z digest=sha256:ae655fdf27d1de903cd42a7a153f3f3bc1a822f32de1b2287afebfa97afb8785

Observation 8de97828-6ac3-4f9c-ad77-8a44bb69ece8 · outbound

This paper cites Mialon, C.

Scaling Test-time Compute for LLM Agents Mialon, C

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:49.315918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:41:48.368652Z digest=sha256:b255907813211d90491b861ab4fe09da84420a48c89106ff408aa61b6d5ffad8

Observation f4ae9f3d-3792-4c18-b3d9-f7059aff5258 · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

Scaling Test-time Compute for LLM Agents ToolRL: Reward is All Tool Learning Needs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.372875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.372875Z digest=sha256:bf7483770ddd70feb4eecd07a172df14f97ab0726417167e45a02f16eab18456

Observation 30d0d066-64fe-407f-b1f6-1820f5f7e9c8 · outbound

This paper cites repository.

Scaling Test-time Compute for LLM Agents repository

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:49.300615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:41:48.376819Z digest=sha256:55c261259695e14d483e6a96e8fb80a0381a2b8cbde8c621086af390b4eaf849

Observation 7d703542-bcf2-4533-851b-c7559cc76899 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Scaling Test-time Compute for LLM Agents Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.380796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.380796Z digest=sha256:bcad1fbacb3d970dec8037468105cf3a19acee64c1f2726a237fee6232a8d236

Observation 84872660-fed1-423e-b193-bf34ebd254ca · outbound

This paper cites an unresolved cited work.

Scaling Test-time Compute for LLM Agents Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.384860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.384860Z digest=sha256:feda1020d7bb34b4417b9420ac1a673c9fe088ce2d5286d802f6e1d1fdad27bb

Observation 929e9e6b-b2dc-4f8c-90d0-fff4808e710f · outbound

This paper cites A Comparative Study on Reasoning Patterns of OpenAI's o1 Model.

Scaling Test-time Compute for LLM Agents A Comparative Study on Reasoning Patterns of OpenAI's o1 Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.388696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.388696Z digest=sha256:4e6a1b2f13c21f367cb53fd3c844497026524125783ebad5ea1df8e0ed56cd0a

Observation dc793ea2-6375-4da3-84d0-54abef17e3a4 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

Scaling Test-time Compute for LLM Agents Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.397018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.397018Z digest=sha256:ecb2a9b6a67f74bb2ca677a66990334415e1be1b80016dd89081d0c7c85ff4da

Observation 2bc32b7b-4c2f-49be-a1ba-f3b1bd33099b · outbound

This paper cites OS-Copilot: Towards Generalist Computer Agents with Self-Improvement.

Scaling Test-time Compute for LLM Agents OS-Copilot: Towards Generalist Computer Agents with Self-Improvement

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.400792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.400792Z digest=sha256:3db5c935bf6c7fdfc3b42321d98c1d57b5fc9a78a9f17d43982262f0d96dbcb8

Observation 97f62b28-d998-499f-aa58-9eb0923c1a72 · outbound

This paper cites Self-rewarding correction for mathematical reasoning.

Scaling Test-time Compute for LLM Agents Self-rewarding correction for mathematical reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.404875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.404875Z digest=sha256:5c98202bc6dac7172cea1b0ff186e1670f8f1d1e7920e747863fafe819bb755a

Observation 60284b91-56cd-4e56-bbea-0ef459298217 · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

Scaling Test-time Compute for LLM Agents WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.409072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.409072Z digest=sha256:cd3074c164b0c5eba37012f4c365116bb2927e457c4298d33cfe7643f22ee21c

Observation 8d6593bd-d119-42e1-9cf7-7f035bae4949 · outbound

This paper cites RecurrentGPT: Interactive Generation of (Arbitrarily) Long Text.

Scaling Test-time Compute for LLM Agents RecurrentGPT: Interactive Generation of (Arbitrarily) Long Text

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.413118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.413118Z digest=sha256:7af9ec5568d47b3a1957222353f7a6ca11d0f0ec538401ee0c80f9434ab10ca6

Observation 524c4b4f-084c-431f-bb09-1a10a6336551 · outbound

This paper cites Agents: An Open-source Framework for Autonomous Language Agents.

Scaling Test-time Compute for LLM Agents Agents: An Open-source Framework for Autonomous Language Agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.417200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.417200Z digest=sha256:709ec22855143619fd01c9b4e13f641416d3589b7634beb8f756908848b4b4a9

Observation 752036d6-2d75-4ae0-8470-afc36563930b · outbound

This paper cites Symbolic Learning Enables Self-Evolving Agents.

Scaling Test-time Compute for LLM Agents Symbolic Learning Enables Self-Evolving Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:48.421514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:48.421514Z digest=sha256:518f7121667310baace5de6a044294da19eba880fde768750f4e7580e4f3eda2

Pith citing papers

Observation c74b1a72-ace6-4ffb-8732-1be03af50989 · inbound

SSRL: Self-Search Reinforcement Learning cites this paper.

SSRL: Self-Search Reinforcement Learning Scaling Test-time Compute for LLM Agents

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T20:17:13.855017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:17:13.855017Z digest=sha256:a6220c9b69f3a65e74040508638d1338728709549237e1a8b0cc417bd8cd2001

Observation 4bc80da3-53d8-4761-b32d-2fd4cf67c913 · inbound

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL cites this paper.

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL Scaling Test-time Compute for LLM Agents

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-05T23:57:37.738674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:57:37.738674Z digest=sha256:b2931351f364422a3eed9e67225298bd41f84d50e560124e6a99841428a16a4b

Observation e3642bd2-4c47-4b87-893b-f479cf8520fd · inbound

Evaluation-driven Scaling for Scientific Discovery cites this paper.

Evaluation-driven Scaling for Scientific Discovery Scaling Test-time Compute for LLM Agents

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:26:05.399233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T03:39:52.204043Z digest=sha256:ff65c84b1320b06203d0ec0b3cab6ba549cd96e5a63afdeb39461cda0c1eb6f7

Observation 782feeaa-f415-49e3-a1e3-c9b58861e895 · inbound

Inference-Time Budget Control for LLM Search Agents cites this paper.

Inference-Time Budget Control for LLM Search Agents Scaling Test-time Compute for LLM Agents

Reference 80

Resolution
malformed identifier
arxiv_id, observed 2026-05-11T19:31:07.932732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T11:51:02.872129Z digest=sha256:a90df62efda9d76f548100765a255bde864a676aa85e3281b545c65e4f5eb450

Observation 4aea3837-ef7b-40e2-a7d4-23fcc92696ec · inbound

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability cites this paper.

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability Scaling Test-time Compute for LLM Agents

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:06:19.459616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T03:03:05.715652Z digest=sha256:9d8958bf066adb67c57ddcf2f0d9763a601f471366ef8a370e6be24aa7b90b8f

Observation aece34da-0f04-4b99-9b93-dce8ded68d13 · inbound

The Granularity Mismatch in Agent Security: Argument-Level Provenance Solves Enforcement and Isolates the LLM Reasoning Bottleneck cites this paper.

The Granularity Mismatch in Agent Security: Argument-Level Provenance Solves Enforcement and Isolates the LLM Reasoning Bottleneck Scaling Test-time Compute for LLM Agents

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:32:02.859506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T01:31:13.100257Z digest=sha256:0e5ca81c2181786c12fd0d15a43eb19f7b9b211c398159e64568381b647233e3

Observation 02bd3d7b-4e54-42c7-bf88-5668d8ebdc51 · inbound

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility cites this paper.

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility Scaling Test-time Compute for LLM Agents

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:39:47.911666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T05:35:09.705532Z digest=sha256:68a47b236c15090f812fde3a76c79be3e8e4866bfc29812d8dc49fafc18e6d63

Observation 5640f2b7-5067-48bb-87a6-90f722a0c543 · inbound

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents cites this paper.

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents Scaling Test-time Compute for LLM Agents

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T22:57:50.188959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T22:55:28.637039Z digest=sha256:c00f75b0beee981fed37c3787d37b8b94db355f97e4c1bd2e4a7e687627f83cb

Observation cd3253b7-eea2-4e30-9d1b-8a9bfee353ef · inbound

ExComm: Exploration-Stage Communication for Error-Resilient Agentic Test-Time Scaling cites this paper.

ExComm: Exploration-Stage Communication for Error-Resilient Agentic Test-Time Scaling Scaling Test-time Compute for LLM Agents

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:54:38.586904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T05:54:12.458110Z digest=sha256:38e046a5deca0a5bf723f2ff2314ba334cc66a74a1b334f3c220eeb6e18f3f4d

Observation 9847c3d3-f2bf-4878-bfff-f2fbe9fefaa5 · inbound

IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents cites this paper.

IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents Scaling Test-time Compute for LLM Agents

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:24:40.468558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T06:21:56.971204Z digest=sha256:e00236292ae08531c1d1b5a4a4c930a38afb5b4300cbb8e6bd7ab7b560778a81

Observation 9e91e587-47b1-422f-9eaf-0e96bb217707 · inbound

CurveShift: Is Agent Progress Scalar? Separating Level from Shape cites this paper.

CurveShift: Is Agent Progress Scalar? Separating Level from Shape Scaling Test-time Compute for LLM Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T00:43:21.854604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:43:21.854604Z digest=sha256:758eb1b3b245f277276dfa9ab7a20edacece43a89e5963c88622037f7822fbc1

Observation 7921223b-3ee0-4633-9e0a-af2fe5e9b1f8 · inbound

Cognitive Demand Steering for Adaptive Meta-Reasoning in Large Language Models cites this paper.

Cognitive Demand Steering for Adaptive Meta-Reasoning in Large Language Models Scaling Test-time Compute for LLM Agents

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T00:24:22.837114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:24:22.837114Z digest=sha256:121fbc9840796d458bc47c5f050174046119bbecd46b4c01fb11a3319d13e1b1

Observation 41ae6c14-ffbc-4112-9389-4ca2b5dde8ee · inbound

CRISP: Critical Step Perception for Training Efficient Deep Search Agents cites this paper.

CRISP: Critical Step Perception for Training Efficient Deep Search Agents Scaling Test-time Compute for LLM Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T19:21:37.747063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:21:37.747063Z digest=sha256:b8b82bc310b0a6cbbff26fc8c3be7bcd51d31a006656229ad6b8efcfb919dbd6

Observation e883c10c-9899-4482-8c63-faedb87f389e · inbound

CRISP: Critical Step Perception for Training Efficient Deep Search Agents cites this paper.

CRISP: Critical Step Perception for Training Efficient Deep Search Agents Scaling Test-time Compute for LLM Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:41.203958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:41.203958Z digest=sha256:3caeea28d2e6db01f46188a75c4054f5337aadffcdd61fed91e2baf631ac135e

Observation ddbf3e45-f320-4df8-b7ba-3f42e3441114 · inbound

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds cites this paper.

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds Scaling Test-time Compute for LLM Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T04:26:29.391821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:26:29.391821Z digest=sha256:427566e42e2f76694152b78cd7341310cbec9b4d1805e35acada24761eb97b5a

Observation 3e914543-90f7-4f96-bbdc-5af9c64d2bb7 · inbound

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution cites this paper.

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution Scaling Test-time Compute for LLM Agents

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:42.130163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T10:20:42.130163Z digest=sha256:7c54d0e1ea62627ab83bf7f0981c15fccd1ba872ea7a7ce6a22c7d45ebfc6e01