Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2312.01072.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:52:33.490521Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T17:40:00.280883Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1cf9356d-aaa7-414f-90fb-033112d371c6 · inbound
Humans Coexist, So Must Embodied Artificial Agents A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 112
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4199dbf-a63e-43ac-abaf-b87a4627bfe1 · inbound
Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 2011
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b1898ce-bb72-4cd4-a68f-97cae96ce57d · inbound
TextAtari: 100K Frames Game Playing with Language Agents A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d113656-ccc7-4533-8eef-555599ea29de · inbound
First Return, Entropy-Eliciting Explore A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 649d76d9-202b-4770-bca8-c573cdc4b89d · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 2007
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5718411b-dbfe-4b54-a177-6b9c5b198636 · inbound
Deep Reinforcement Learning for Dynamic Origin-Destination Matrix Estimation in Microscopic Traffic Simulations Considering Credit Assignment A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b211e744-065c-4019-a2ea-774e7c3d15be · inbound
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61e63b3e-dccd-4650-9cca-922aa3884b32 · inbound
Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dc88ae64-b444-48d1-a77b-4748766ddacb · inbound
From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 082c771b-189a-4aaa-8a2b-f1d86300bc87 · inbound
Moira: Language-driven Hierarchical Reinforcement Learning for Pair Trading A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e9c8f779-438d-48d6-9b79-b95053d0e646 · inbound
Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 53a38aa9-158d-46e4-abea-047a8bc05ff3 · inbound
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d757079b-5053-4254-a5e2-deba08a82383 · inbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8001aa6d-80f9-444b-8b78-e69cf61bdc08 · inbound
GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0f2e0b28-0088-4b84-b85e-54dba3336231 · inbound
When Denser Credit Is Not Enough: Evidence-Calibrated Policy Optimization for Long-Horizon LLM Agent Training A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 48565c02-186a-4af8-aa6b-116c7c86ddfb · inbound
Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 90b1c0b3-c1bb-41ba-b5ed-2b4552bfd87e · inbound
World Value Models for Robotic Manipulation A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 71116b30-7350-4d8d-9dca-5b46c4569bae · inbound
Attribution Markets: A Fisher-Market Formulation for Fractional Credit Assignment Between Planned Tasks and Performed Actions A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 876c39cb-1e01-4873-8117-0980b79e7276 · inbound
Learning to Prepare Molecular Ground States with Transformer Models A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a1e9ed0-3451-4c16-857d-22bd399370dc · inbound
Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.