Pith. sign in

Paper Citation Record · LEDGER

Reinforcing Language Agents via Policy Optimization with Action Decomposition

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2405.15821.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.15821 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:37:47.590688Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T01:48:28.550913Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8029fb17-4eaa-44da-a0b9-37a8b23ba8c9 · inbound

ProgRM: Build Better GUI Agents with Progress Rewards cites this paper.

ProgRM: Build Better GUI Agents with Progress Rewards Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:47.590688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:47.590688Z digest=sha256:1363d784923d1c80bd7cf7914aa3035cbf03638cb485aa24abe2ff476894cc00

Observation c6e8579a-c8a1-4ccf-b5e2-58311378a9dd · inbound

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models cites this paper.

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:30:58.541659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:09:36.341574Z digest=sha256:bf11b37a3a6ebd8fd1e5b75ad490765df54fac405817f8120ca25174f0f1b175

Observation ee7f9348-e1a5-4dd5-b9ff-df2570a48912 · inbound

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy cites this paper.

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:48:28.553279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T01:46:24.724553Z digest=sha256:68bb5feb5f32e5ffd9bfe431e6cb452eca8edbd3cc62e51303e8864717191568

Observation 053a9567-4e62-4d7c-8e96-0d99d0544034 · inbound

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent cites this paper.

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent Reinforcing Language Agents via Policy Optimization with Action Decomposition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T00:23:00.148968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:23:00.148968Z digest=sha256:6aedec9f8aaa441a62e63847b0030d4953eb790888789211707ef30de86ce099