Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:55:14.710636Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2411.19809.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:55:14.710636Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:35:50.981161Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T13:35:52.227123Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9b7c3d74-eaa0-4e75-8e10-735f684148b2 · outbound
Q-learning-based Model-free Safety Filter Rapid Locomotion via Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7ee1373-3bf3-48f5-ad14-d924720bef00 · outbound
Q-learning-based Model-free Safety Filter Deep Reinforcement Learning for Robotic Manipulation with Asynchronous Off-Policy Updates
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46372c24-a799-4920-9c07-ec66d651a4f1 · outbound
Q-learning-based Model-free Safety Filter Barrier- certified adaptive reinforcement learning with applications to brushbot navigation,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5471a02c-9c3a-4a40-bc87-74051ea2fdca · outbound
Q-learning-based Model-free Safety Filter Safe reinforcement learning using robust control barrier functions,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c4b7b39d-0e86-4723-8098-50dd7e88249b · outbound
Q-learning-based Model-free Safety Filter End-to-End Safe Reinforcement Learning through Barrier Functions for Safety-Critical Continuous Control Tasks
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 858acfdf-2c74-47c8-94f9-2d8d33c1a165 · outbound
Q-learning-based Model-free Safety Filter Probabilistic model predictive safety certification for learning-based control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9327079e-9592-4a50-8d16-f357ece1aa49 · outbound
Q-learning-based Model-free Safety Filter A General Safety Framework for Learning-Based Control in Uncertain Robotic Systems
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e3901c7-0b86-4289-b347-5af2be9e7cab · outbound
Q-learning-based Model-free Safety Filter Agile But Safe: Learning Collision-Free High-Speed Legged Locomotion
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac379a51-0a69-4ca4-bb16-2d0bc11f4b42 · outbound
Q-learning-based Model-free Safety Filter Constrained policy optimization,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 718ed895-76f5-4502-98ff-c3d91556111f · outbound
Q-learning-based Model-free Safety Filter Value functions are control barrier functions: Verification of safe policies using control theory
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e06bc79a-b879-4f01-b3c3-6ac56719ca4a · outbound
Q-learning-based Model-free Safety Filter Bridging hamilton-jacobi safety analysis and reinforcement learning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 84a06bb5-228b-432c-8639-997a0bf8979a · outbound
Q-learning-based Model-free Safety Filter Altman, Constrained Markov Decision Processes
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 18533d8f-d73e-4b17-bec4-aa740caec948 · outbound
Q-learning-based Model-free Safety Filter Adaptive dynamic programming for nonaffine nonlinear optimal control problem with state constraints,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffada0da-e970-4f27-a5f1-4c090bbfc393 · outbound
Q-learning-based Model-free Safety Filter Reward Constrained Policy Optimization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42fabd6b-c9c7-491f-9087-fd471dbf8f9d · outbound
Q-learning-based Model-free Safety Filter Benchmarking safe exploration in deep reinforcement learning,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0cb7765-bfe3-436d-9425-50d9c8bfde69 · outbound
Q-learning-based Model-free Safety Filter Integrated Decision and Control: Towards Interpretable and Computationally Efficient Driving Intelligence
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 5fc2dc98-115d-4ab1-8287-2facf86f4007 · outbound
Q-learning-based Model-free Safety Filter Learning to be Safe: Deep RL with a Safety Critic
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c01971df-14d8-4442-b45e-ff6d2f050f3a · outbound
Q-learning-based Model-free Safety Filter Re- covery rl: Safe reinforcement learning with learned recovery zones,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e3dccc95-03b9-47d7-b388-d61c78a5f02c · outbound
Q-learning-based Model-free Safety Filter Safety Optimized Reinforcement Learning via Multi-Objective Policy Optimization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cecd247b-e3da-4f8c-842e-df1afbc8b5bc · outbound
Q-learning-based Model-free Safety Filter Safe Reinforcement Learning by Imagining the Near Future
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 70fa88f3-eb34-4674-ab67-d16b00b05eac · outbound
Q-learning-based Model-free Safety Filter Learning safety critics via a non-contractive binary bellman operator
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2c8fc713-064f-4d8f-b960-2555d1c80c68 · outbound
Q-learning-based Model-free Safety Filter Constrained Policy Optimization via Bayesian World Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d26d5cb9-9839-4e09-af70-96ef11405211 · outbound
Q-learning-based Model-free Safety Filter Safe Continuous Control with Constrained Model-Based Policy Optimization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0b7932dc-7da0-4293-a712-d0e32fa210a3 · outbound
Q-learning-based Model-free Safety Filter Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09c2efcb-a79d-47fa-b98e-8ed4799321c8 · outbound
Q-learning-based Model-free Safety Filter Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be8e5220-3762-43f4-8493-954afa1db199 · outbound
Q-learning-based Model-free Safety Filter Playing Atari with Deep Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a6b6002-6338-43b3-ab35-bd2b7df9dfed · outbound
Q-learning-based Model-free Safety Filter Proximal Policy Optimization Algorithms
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9e2ffd1-65e2-486f-b06a-2db503f9bb38 · outbound
Q-learning-based Model-free Safety Filter Continuous control with deep reinforcement learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fc480f0-a8f1-4e36-a9fc-4cb1f5df8264 · outbound
Q-learning-based Model-free Safety Filter Addressing Function Approximation Error in Actor-Critic Methods
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 266b26f6-922f-4181-a350-432b9eff8ec6 · outbound
Q-learning-based Model-free Safety Filter Optimal control for a shape memory alloy actuated soft digit using iterative learning control,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d1474a7f-e3fa-4f85-b16c-05acae6b17cd · outbound
Q-learning-based Model-free Safety Filter Model-Based Reinforcement Learning for Atari
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57d2901b-701d-40f7-88a2-c35fb3078c04 · outbound
Q-learning-based Model-free Safety Filter Projection-Based Constrained Policy Optimization
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fefe6960-2dc7-4e1d-af49-7c0ed08d36d4 · outbound
Q-learning-based Model-free Safety Filter Safe Reinforcement Learning Using Robust Control Barrier Functions
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85f2e082-e725-4692-8836-3ec34871f147 · inbound
Verifiable Safety Q-Filters via Hamilton-Jacobi Reachability and Multiplicative Q-Networks Q-learning-based Model-free Safety Filter
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.