Pith. sign in

Paper Citation Record · LEDGER

Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2308.02151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.02151 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:07:58.380043Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:58.626281Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f8659ef3-82dc-41a9-a27b-7aafb6eb37c1 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 234

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.056161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:0f281db10f2783d09613f04b08babd696670439054b1cd25890b8a2d044ab2d7

Observation c960f693-2689-47f1-85d4-6dd2a9dec05e · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.992190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:e5b84cb8222bbe4a40977906190ad00ae289315b6bed3baab06683ebb3d6c6ea

Observation 4b153e26-3749-4212-8f15-760e5d54fba9 · inbound

Refining Answer Distributions for Improved Large Language Model Reasoning cites this paper.

Refining Answer Distributions for Improved Large Language Model Reasoning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T13:20:59.595233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:20:59.595233Z digest=sha256:029c791737da1c07da59522f891accba60117fae48f16b7fe4e5ae50167cc15f

Observation 18dba6ac-c970-405b-9ed8-23e724da7a4f · inbound

Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches cites this paper.

Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:12.416838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:12.416838Z digest=sha256:f5fb62dfa312227999fb9db3fbcaa859cd2b992c0e88518d705a27d5fd493d67

Observation 6682630d-db10-425d-8e0a-61d2a7e13492 · inbound

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs cites this paper.

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:05:09.813162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T11:05:09.588491Z digest=sha256:c2dcbbd8465bb4b4f05a2d047194df2c3f0e5eb4ad05ea5c57068443222f252f

Observation 9ccbaf35-230a-48b0-b377-073eb13235d9 · inbound

A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms cites this paper.

A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T11:07:58.380043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:07:58.380043Z digest=sha256:321e4e135546333ab17b751947058b66b78b9c752e3eeae0acb31eeef7c11242

Observation 5078c01b-bbcd-4c12-a041-e024a5cc104b · inbound

Model Performance-Guided Evaluation Data Selection for Effective Prompt Optimization cites this paper.

Model Performance-Guided Evaluation Data Selection for Effective Prompt Optimization Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T21:08:46.705488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:08:46.705488Z digest=sha256:3397219ad43fb31581305db2226edea96b2e72682fc7f7a017563aa466814de2

Observation e1d86424-61ad-4b6a-ac8c-1e34c4195ff6 · inbound

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection cites this paper.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.635704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.635704Z digest=sha256:52b63ef01fc8fd3529fbaa50b3dc60e90f99a0ce7a4265c5d40446d8bf34cbfc

Observation 498ff026-fe3f-4f37-bd26-0a457ef52ece · inbound

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team cites this paper.

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:14.981510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:24:14.981510Z digest=sha256:3b790b4111d6c1447e347abde146324bd7fc18997883f8b1ade4c4654748b78a

Observation c23a2c9e-f99b-4001-8865-9b3490f4bf92 · inbound

Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager cites this paper.

Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:35:12.612649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:35:12.612649Z digest=sha256:be99dd764c7ffaef0ecac2bd34fc108ba3f358c476a74d7e1bfc47032d3ad7df

Observation d97be4d3-3c47-4cd2-b2ba-ae416b0d23cf · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:08.681011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-08T10:23:52.522238Z digest=sha256:1cb2c1de5abb05517f2de3d3b11c6a41448ff22ba0d983a6b35a932560bbab26

Observation fd45479c-89ef-47c3-8787-9f632c461736 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:00:56.846868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-11T02:00:00.663355Z digest=sha256:c499696a4b264793a4143b2930182822a5359c56db1476ce2cf45356b4afe154

Observation 2257a956-2335-440d-96e2-ff68c3f54f44 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:17:28.208672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-13T07:17:13.708752Z digest=sha256:0d2f677997607aa20de5edbc4bb626a27a77eb26e6fb2235d9b2686860b17f6e

Observation e9983242-aa4a-4fd2-b415-f426ffd395d4 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.093072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:9d1ebd2621b6752c7e6c31d59da341887af54160bf42c3cca8d75354ab96c87d

Observation 919c1440-c2f1-490e-b7ce-0b418b1e4dae · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:14.904430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:e005a95b0777e546bd6044af6fae2d88d7fd7c95f0dcc8f48bae2e1db2a4503e

Observation 20f42e94-cf13-40f8-96f7-4ab6f5e594db · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.333934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:0d4404c2e511342a8b8c032c72297c7e81e7b85b2429220c313daabf67098fb0

Observation b931cb18-5ab5-47e2-b9d4-e2fa4fe981fe · inbound

Training Language Agents to Learn from Experience cites this paper.

Training Language Agents to Learn from Experience Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:14:02.615242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T07:11:09.642275Z digest=sha256:c5ac21284f367878db50686c48174b642c26f0b49a3a1b3be3bf490583e1c8b8

Observation 84dc25dd-01a6-4f6c-a937-6bf2731214a6 · inbound

EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation cites this paper.

EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:28:58.628130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T00:34:47.530464Z digest=sha256:cad6a304bbe383f59054397ea1a20d5f814a542c294250b498690d93216cd405

Observation 2d5a709f-7554-4495-9096-3cfa72d7bed2 · inbound

S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF cites this paper.

S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T13:58:42.937172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:58:42.937172Z digest=sha256:46600c71203a9055214b1da9afb7e0d64263b219b764568aebf952e7e6d173bd