Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2303.17396.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-31T11:28:31.931515Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 5edb3c11-8fc5-4400-b5af-2af1766ec643 · inbound
Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea6b5855-b0e0-4111-a622-dda54e4e03f6 · inbound
Reinforcement Learning with Action Chunking Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d2898d64-945d-4361-81bc-da44e88c9c50 · inbound
COOPO: Cyclic Offline-Online Policy Optimization Algorithm Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f41da937-d31d-4bf2-a611-c38b899f92fa · inbound
Improving Robotic Generalist Policies via Flow Reversal Steering Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7fa6845a-7bf4-4fe6-a75b-8ce07cb021ef · inbound
Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca81e43d-d058-45b3-b7a0-2de550c049aa · inbound
Robo-ValueRL: Reliable Value Estimation for Offline-to-Online Reinforcement Learning Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87b90e74-70a7-421b-87d2-1abcf9db6afc · inbound
Evaluating Fuzz Testing for Reinforcement Learning Agents Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.