Pith. sign in

Paper Citation Record · LEDGER

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 6 inbound Pith citation observations for arXiv:2507.01489.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.01489 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:55:09.996272Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:55.496339Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T13:53:28.611729Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2980a235-2774-4187-bc1b-f97f589fd2ef · outbound

This paper cites Brown, B.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Brown, B

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:55:10.276010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:55:08.809952Z digest=sha256:d8cfc9551261adbe2b3934d3506f0469494cbc1d6ae70465fe88edbbdae92cdb

Observation 741dd78a-e701-4532-96fa-7b09cd18633c · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.044949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.044949Z digest=sha256:13fec45005862ae3e41b9703c1dde90a60e59dc1beb120bcdee3f2c220f3a234

Observation b56633a1-1bb9-4945-9047-497a5ffb5378 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.251681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.251681Z digest=sha256:eaf30b4fee4b11b87c815f3ff895c5c1db8c941c8852e766b6c06c06c8443823

Observation cd883f29-6f39-45e4-a322-8bb6646e2d57 · outbound

This paper cites Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.340912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.340912Z digest=sha256:34b017d762ff8cf514556e60bbe94e001a021752b6881f0a1455a09c51aa9cd8

Observation 579394ad-e4ce-4fc6-8220-841b831c109b · outbound

This paper cites Measuring and Narrowing the Compositionality Gap in Language Models.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Measuring and Narrowing the Compositionality Gap in Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.438805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.438805Z digest=sha256:d2d74e9f30a33f7bfd9e995b00fbd8fa99c7b637108ac63abf1e1033803a8c58

Observation e4058c2a-53ed-4ea3-a043-9ebed6af1788 · outbound

This paper cites Qwen2.5 Technical Report.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Qwen2.5 Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.563729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.563729Z digest=sha256:35d3257af627bd1c301a46ca1878a74bcc13775a4066c418a1cb8694a3c818f3

Observation 65766e1d-9f9c-413a-b2a4-8a57bdb89f94 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.825758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.825758Z digest=sha256:115b3661fb6f94f5c95940aa050ebc2479b8b66c69a392b52277406a03d98055

Observation 5a78b56a-609d-4642-bbdd-b527cbff4f49 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning ReAct: Synergizing Reasoning and Acting in Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.879680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.879680Z digest=sha256:06970cac180d9a182098926ba8ee3a116b5d3e2c959a57235a0b54b56a1decbc

Observation cc12c887-380f-4ca6-a429-67fadbedd70f · outbound

This paper cites OpenResearcher: Unleashing AI for Accelerated Scientific Research.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning OpenResearcher: Unleashing AI for Accelerated Scientific Research

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.948075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.948075Z digest=sha256:f9eb766d48bebac143f9ef67c8bc61c2e94b835355d5561839e3253128072c6f

Observation 3779de78-c0be-4671-ac6a-69f58a4549d4 · outbound

This paper cites DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.996272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.996272Z digest=sha256:6cbb92d44617e4b5058e21e152fdd95e94b4858ba5fbd6a3316d65a85ec8ad22

Observation bb878f04-f70a-467a-a495-7e8d242f156b · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.108495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.108495Z digest=sha256:b7eabfd1615dd28c90654578a5777f8361d2c98a68318e0cd567ce774d64520b

Observation e7748746-bc8d-4743-82b9-64a1f7610247 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.770719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.770719Z digest=sha256:d3d3d1a05cafdc749e801bc8ec13e6e53208dc17effa52027cbd162450c19e55

Observation cc514193-ee30-4d29-9e02-81507033a662 · outbound

This paper cites GPT-4o System Card.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning GPT-4o System Card

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.165242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.165242Z digest=sha256:1e6a8bcd3b46b94c957c65fc2cd6c9e136a8a2f78a0fe3429c5416bd556ae79c

Observation 44c22c11-5e34-44c1-a8b0-410e413b38c8 · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:09.695260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:09.695260Z digest=sha256:ad34a32d08d4103dcabe72d538157421396647778c05aac99ccc03414f5ca07c

Observation 4b64dbb8-5014-4415-9e30-908677aa8367 · outbound

This paper cites Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use.

Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:08.936392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:08.936392Z digest=sha256:5389663d1582994d4d72f4c2ce3b84b3bfdb96fe10a64c840ea4d5dfee3b5026

Pith citing papers

Observation 88b47baf-aff7-4e85-bc8e-def515eb4e0b · inbound

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory cites this paper.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Reference 228

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:15.635577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T23:13:15.016486Z digest=sha256:6afc347a1553cf63c49f5caf4bb1a32add5c08fa2c45e98e3d0d61a73a680ea3

Observation 56960e50-3558-4410-8bca-ed3bd3392f08 · inbound

GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation cites this paper.

GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:47:26.597990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T06:44:30.069444Z digest=sha256:fb505b6f508a713e2ef11886f77cd54dbb2dda9d4bfdd128bc527388d82b995f

Observation 91881407-ef2d-4594-a886-d4bd1467fc30 · inbound

GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation cites this paper.

GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:59:48.707183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:59:23.949124Z digest=sha256:bb04780a6ec37f89807e74739b2a31e2f94e4384cb54f7983d6f5e8f0e491d59

Observation 2d126fbc-da39-44de-b05f-8ef1dc014e1f · inbound

IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents cites this paper.

IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:24:40.531043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T06:21:56.971204Z digest=sha256:bd99e1536bafb0b1e9d392e0a91a31097a5a1240699825ffc9ffcd82a18feb06

Observation 3ae45dc3-c1ce-487e-9d3b-aa30ec874642 · inbound

Knowing When to Ask: Segment-Level Credit Assignment for LLM Tool Use cites this paper.

Knowing When to Ask: Segment-Level Credit Assignment for LLM Tool Use Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:53:28.613407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:49:58.677299Z digest=sha256:d88df6ab3611ab59a9b835d291bdfe219a18f46ed6e0d1ada03e9766956e8644

Observation 4efbff9a-c058-4e2b-9631-ae4d821c1538 · inbound

OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems cites this paper.

OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:55.496339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:55:55.496339Z digest=sha256:3812f4fc3f63a2e208f9443526421e69dae0bcd2f84e818b9408eb4f5c52893d