Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2002.02829.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-30T14:49:29.137552Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T06:11:22.839895Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9bdc9302-e4c8-40e7-962e-0a641dce73a8 · inbound
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9bd90051-df4b-4f1c-a682-047eff9bd130 · inbound
Autonomous Transition State Search with Soft Actor-Critic Reinforcement Learning Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e540d67b-2188-418b-b546-5bc8c3045a16 · inbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a40db98-bb95-44ea-b329-24d0d2ff932c · inbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.