Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:27:44.056272Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2607.27203.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:27:44.056272Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
64 of 64 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fc8d6680-afd6-4aa9-abff-3fbe039a3cd8 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Efficient Online Reinforcement Learning with Offline Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 435b5e08-e59e-4b49-8514-ea7445d15ea1 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49b28c9e-8205-4ea0-b167-2c60e716a353 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Reinforcement Learning via Implicit Imitation Guidance
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3de1e9ba-ad4e-4dc2-8962-641f2c84653a · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? EXPO: Stable Reinforcement Learning with Expressive Policies
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b753a510-00a3-4072-b11f-db5da5f4506f · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4f1dd5b7-39d8-4d00-bd48-d23ff9b982ad · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Tql: Scaling q-functions with transformers by preventing attention collapse, 2026 b
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfb8630c-ab78-40c1-9a6e-8839ca96ed43 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Value Flows
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ba61b2f-ad5d-4c3f-90f1-50f313f9c14c · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? A Minimalist Approach to Offline Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db5f9166-ea8d-42c8-b127-c020905a6e98 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Off-Policy Deep Reinforcement Learning without Exploration
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e76a3389-5aca-4375-83dd-25a19774da10 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30ce3aad-ceb7-4613-88d2-e5bc3e2ef94e · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4ef21e19-4a49-49a0-bd4e-28aec6a260ac · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Residual Reinforcement Learning for Robot Control
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8732657e-48fb-40c0-8468-94f37b81e5f2 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b1a3d1c-7b99-4f72-a8d7-3ff78b3bd4d5 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? What Matters in Learning from Offline Human Demonstrations for Robot Manipulation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58830bf1-7272-4d8e-91da-9c5b9b5759ea · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc29c0e7-e20b-4a77-80af-d7e93fb93bb6 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ed5d8e6-24f2-4902-a4e4-8c77af3b4e91 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Training language models to follow instructions with human feedback
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59243467-6b8b-404f-be6d-410361a58dae · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? OGBench: Benchmarking Offline Goal-Conditioned RL
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56944067-1641-4a97-b8df-a49a69282f82 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Flow Q-Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f3bcbff-1be5-4146-9a4e-5b34e7f19672 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d59be1e-05dd-4196-8f89-cffe8033cd6b · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Diffusion Policy Policy Optimization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c59d1e37-14ea-4bb7-b8ce-d3f0f8f81e93 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Learning from demonstration
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 32783b5f-8530-4ddd-ac87-dceba0a00845 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Proximal Policy Optimization Algorithms
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63bc5cb3-043d-4cd5-89ed-347195cd729c · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Residual Policy Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb0d603a-dfea-4cfe-a598-f91434d33dae · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Gemini Robotics 1.5: Pushing the Frontier of Generalist Robots with Advanced Embodied Reasoning, Thinking, and Motion Transfer
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bb83959-8101-4961-9f60-cad24838613f · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Jump-start reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5cb8093f-1c55-45d5-b11b-e15ab67eeaa8 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9a10551-8f68-4ff2-845b-e4b2c0c6343a · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Posterior behavioral cloning: Pretraining bc policies for efficient rl finetuning, 2025
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 927ceae2-6df2-4eb7-b7a2-05f2ce14d0a3 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Hybrid policy optimization from imperfect demonstrations
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 94deceb8-0c11-4433-b87a-394c4ee5677b · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Policy Expansion for Bridging Offline-to-Online Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dce039d1-cbdd-4a12-8533-f5f4580e6e34 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1066cbd0-4f29-44c3-9c82-b26fb9b31e4c · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2024 , eprint=
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7835ab6d-fdf0-4e5a-af30-89b276e89431 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Advances in Neural Information Processing Systems , volume=
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 78f74e7c-1372-412e-805e-151b93895833 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ff6ca7d4-31ee-4b3f-86c9-dae97b4d45b1 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Proceedings of the 40th International Conference on Machine Learning , pages =
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39bfc212-b23c-4369-ad9b-b4b6baa6585e · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Offline Retraining for Online RL: Decoupled Policy Learning to Mitigate Exploration Bias
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8694ecfc-f619-4c79-bc33-0cc52e54f1d3 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0d33480c-c323-4ef5-881f-ccfe952540f4 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2018 , eprint=
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e789b70c-4153-4194-9bf6-a7c476d0876e · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d1980857-ee60-4eba-9aa3-c30134053f99 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dcf5a4ef-cd04-41e1-9252-af049d922457 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2023 , eprint=
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2787269-2f10-4c59-a4c2-3dec0ccc7b6f · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2026 , eprint=
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6e3502c1-8b5f-4acf-8f1a-3b3868634f40 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2022 , eprint=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 261c0147-f982-4990-b942-df7184291f2e · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? $\pi^{*}_{0.6}$: a VLA That Learns From Experience
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96736635-cc01-4463-a940-97b4f349a94e · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2026 , eprint=
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e588344d-82c4-4db6-9258-71db439f99c9 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2019 , eprint=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3340c488-d9d8-4969-8f98-e65a4360125b · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2021 , eprint=
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6e818da6-7c3f-4ff9-85f4-d8474843bbfa · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2018 , eprint=
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ff2fc180-aa2b-46f3-ae96-a689cd150192 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2021 , eprint=
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation abc7736b-dbb5-47b9-a59f-a4c8830b68ad · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8e547edc-336b-4161-8db1-bd7993eca08b · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2019 , eprint=
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d034223c-a59b-4923-ac80-a71d0870a309 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2018 , eprint=
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c8d85fc6-f76f-49cd-a26c-0f1333988e7a · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2021 , eprint=
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47dd6549-d3ae-4431-af91-e43b699b93cc · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Imitation Bootstrapped Reinforcement Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f7ed219-f687-47a8-997d-2073b0fb3a45 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2017 , eprint=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2d80ac2-af79-401f-8317-cec902a0e667 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Learning from Demonstration , url =
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a4a62e94-6014-4bc9-8471-5a6f91f827f0 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2018 , eprint=
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c5580547-0659-4f5a-bcce-2ef08c83c0a9 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2024 , eprint=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 388ee01e-ae08-48d0-a270-2a8eb908f0e5 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f4d91b64-1675-4b94-a1a9-8fd81f082c41 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2026 , eprint=
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b68ff781-7502-41e6-912b-c022c4320ba4 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c548dbad-505e-48e7-9df4-4407752df131 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2023 , eprint=
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 494a7ca2-e0ed-46dc-8428-a1cd07b0dfa0 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2023 , eprint=
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 48b63200-b705-4a67-beb3-e8378cfcfaf2 · outbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? 2025 , eprint=
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.