Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2409.09345.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:05:50.935027Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-14T00:26:48.578074Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ac43e7f8-0af6-441c-a216-30d54a835a2d · inbound
Aviary: training language agents on challenging scientific tasks Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 055d0052-2503-4156-b4c9-94ca80654ecb · inbound
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2b80174-0589-444b-9c80-df89f3992d60 · inbound
QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f110b1a-d3c9-49e0-b9a1-cccdc5bde5d2 · inbound
ToolRL: Reward is All Tool Learning Needs Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation abbbeb2b-ce29-407a-843f-5a0d3db5bcaf · inbound
DSADF: Thinking Fast and Slow for Decision Making Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72928ad8-48a3-482c-9f98-3429532b762e · inbound
Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.