Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:51:10.107804Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 3 inbound Pith citation observations for arXiv:2502.05244.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:51:10.107804Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:18:57.677914Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
16 of 16 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 88ea348b-39a9-4d1f-970b-acb033221157 · outbound
Probabilistic Artificial Intelligence Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffca75f1-664b-4557-80c5-38bb985f4dd3 · outbound
Probabilistic Artificial Intelligence A detailed treatment of Doob's theorem
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d73a48d9-9e5a-49bb-b27c-57f90561eb63 · outbound
Probabilistic Artificial Intelligence An overview of gradient descent optimization algorithms
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b50abdd6-bbfb-49c3-8aca-811a9445444a · outbound
Probabilistic Artificial Intelligence Proximal Policy Optimization Algorithms
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c11e398-21fb-4bf5-96b5-15080ba93b8c · outbound
Probabilistic Artificial Intelligence DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a36ca27a-c381-46e4-a85b-a0960fc6b2fa · outbound
Probabilistic Artificial Intelligence Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fccb7db7-c78b-47d5-8a1a-f7d8ff09cc75 · outbound
Probabilistic Artificial Intelligence Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80094c8b-5b20-472e-8415-45ff012c3829 · outbound
Probabilistic Artificial Intelligence OGBench: Benchmarking Offline Goal-Conditioned RL
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceb3ae33-a214-48ea-a9b7-9747da5c0bc8 · outbound
Probabilistic Artificial Intelligence Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a330d4fb-6c83-4110-968d-1628a122c00f · outbound
Probabilistic Artificial Intelligence Improving neural networks by preventing co-adaptation of feature detectors
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106bcbf8-c45c-4402-97f3-b9088c4607a6 · outbound
Probabilistic Artificial Intelligence Humans are not Boltzmann Distributions: Challenges and Opportunities for Modelling Human Feedback and Interaction in Reinforcement Learning
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7e55e571-0262-4fea-a79b-8fd53951eaf5 · outbound
Probabilistic Artificial Intelligence DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 420bb6f7-8e4c-499f-86de-f9b618cfa9dc · outbound
Probabilistic Artificial Intelligence Application of the theory of martingales
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ff63a438-8d3d-4cc2-9371-7cbba5275e38 · outbound
Probabilistic Artificial Intelligence Active Fine-Tuning of Multi-Task Policies
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07804eea-4b68-4e2f-b02a-ef821aa02a34 · outbound
Probabilistic Artificial Intelligence Diffusions hypercontractives
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b1079f1f-26dd-44e4-8b5d-cc184806ec9e · outbound
Probabilistic Artificial Intelligence Soft Actor-Critic Algorithms and Applications
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fb7dec9-0397-4fe8-a38a-64739827875a · inbound
Scalable Bayesian Monte Carlo: fast uncertainty estimation beyond deep ensembles Probabilistic Artificial Intelligence
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc784a62-3329-40ac-83c7-7c6a4c39d286 · inbound
GPU Performance of an Entropy-Stable Discontinuous Galerkin Euler Solver with Non-Conservative Terms Probabilistic Artificial Intelligence
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 531c95fc-ff85-4a56-83d5-365affd244b3 · inbound
Active Learning for Calibrating Entangling Gates via Surrogate-Based Optimization Probabilistic Artificial Intelligence
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.