Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2105.08140.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T08:40:52.852816Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T08:43:15.128113Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cd89333e-74bc-4e73-9bc8-3c96b01c2613 · inbound
Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c4d26f6-1093-42de-b6f9-f4cb7be3b953 · inbound
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1273e49c-7c95-4b3f-84d3-f000ba58e8fb · inbound
UNIQ: Conformal Calibration for Adaptive Conservatism in Offline Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b02f5f93-5f63-4162-94f2-c034f54f4e04 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 176
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e3ed705-f1f9-4f08-9ccb-495cb2c68033 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 177
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a074f1fd-fa65-4776-a518-e65e3bfced2f · inbound
Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b78a919-2db4-40d0-8f21-70201db89ad9 · inbound
Uncertainty-Guided LLM Semantic Augmentation for Heterogeneous Treatment Effect Estimation Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.