Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-30T14:42:38.148642Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2605.24405.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-30T14:42:38.148642Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5da9fbb1-23a1-4d1c-9df8-67cd4e55fdf5 · outbound
Generative OOD-regularized Model-based Policy Optimization Let Offline RL Flow: Training Conservative Agents in the Latent Space of Normalizing Flows
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 815bfa80-503a-4d90-b418-f8c8140a4d57 · outbound
Generative OOD-regularized Model-based Policy Optimization Uncertainty-based offline reinforcement learning with diversified q-ensemble
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1ae7f9f9-2fec-4e66-a50a-ab00a8ea0a59 · outbound
Generative OOD-regularized Model-based Policy Optimization Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 409da6ff-7cb4-4d19-b73b-a914f7ad825d · outbound
Generative OOD-regularized Model-based Policy Optimization Offline RL without off-policy evaluation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 110cdc87-56f5-4778-8bb2-e376dc65e117 · outbound
Generative OOD-regularized Model-based Policy Optimization Schoellig
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 36bda10e-673c-45d9-8f79-c5a040e5eb4e · outbound
Generative OOD-regularized Model-based Policy Optimization Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8480ee01-8411-4721-bfdf-227ad0f7c446 · outbound
Generative OOD-regularized Model-based Policy Optimization Importance weighted autoencoders
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87b23479-b65d-424f-9e2a-99d2e938b07a · outbound
Generative OOD-regularized Model-based Policy Optimization Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a1463f6a-0ca5-4bd2-bfc1-c1dde5bb7794 · outbound
Generative OOD-regularized Model-based Policy Optimization Mag- netic control of tokamak plasmas through deep reinforcement learning.Nature, 602(7897):414– 419, 2022
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 24d483c9-3e2f-470c-8e1b-b7d5c0efe65b · outbound
Generative OOD-regularized Model-based Policy Optimization Density estimation using real nvp
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation af786078-ac63-4022-af98-4786245712e4 · outbound
Generative OOD-regularized Model-based Policy Optimization An image is worth 16x16 words: Transformers for image recognition at scale
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 78686516-463e-48b4-a6d9-67d11bd44adf · outbound
Generative OOD-regularized Model-based Policy Optimization Reinforce- ment learning for precision oncology.Cancers, 13(18):4624, 2021
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 39d684e0-5431-4793-a4a7-9971d538a42d · outbound
Generative OOD-regularized Model-based Policy Optimization D4RL: Datasets for deep data-driven reinforcement learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b980eb7d-833c-46b2-bd15-e9963f6a5b42 · outbound
Generative OOD-regularized Model-based Policy Optimization Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f649f1c9-e03e-426e-a981-2f38c120a5da · outbound
Generative OOD-regularized Model-based Policy Optimization Mean flows for one-step generative modeling
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5625a248-8d3e-493e-b7af-c87c40eac2ac · outbound
Generative OOD-regularized Model-based Policy Optimization Denoising diffusion probabilistic models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fbd8e80c-1fd8-4bc3-bb6b-72085d9702d9 · outbound
Generative OOD-regularized Model-based Policy Optimization Planning with diffusion for flexible behavior synthesis
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 732dfd34-08c6-44bc-bf88-33d46d64abad · outbound
Generative OOD-regularized Model-based Policy Optimization When to trust your model: Model- based policy optimization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aaf15b7c-15b5-4e62-9ac6-55f255c4574f · outbound
Generative OOD-regularized Model-based Policy Optimization Billion-scale similarity search with gpus.IEEE Transactions on Big Data, 7(3):535–547
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e23226c3-83a3-4261-b7f0-c109214678f1 · outbound
Generative OOD-regularized Model-based Policy Optimization Elucidating the design space of diffusion-based generative models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 485d9e79-6ca2-472c-8fbd-b2133e0efdc9 · outbound
Generative OOD-regularized Model-based Policy Optimization Kingma and Max Welling
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 031648cf-0cc8-42f7-b7ae-4295c46a05e6 · outbound
Generative OOD-regularized Model-based Policy Optimization Celi, Omar Badawi, Anthony C
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 415afcb8-c957-426d-89dc-0aee7ba50b2a · outbound
Generative OOD-regularized Model-based Policy Optimization Stabilizing off-policy q-learning via bootstrapping error reduction
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6e63e8b0-fa8d-42e9-ad6f-f610a13b490d · outbound
Generative OOD-regularized Model-based Policy Optimization Conservative q-learning for offline reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9e9c12a0-3b45-43a9-b4ed-33279d2f0698 · outbound
Generative OOD-regularized Model-based Policy Optimization Showing your offline reinforcement learning work: Online evaluation budget matters
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d89c11cb-c5f3-4f10-93a1-63a1a7538f8e · outbound
Generative OOD-regularized Model-based Policy Optimization Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b9d4efb0-dea8-4f77-b433-90cf08fa6671 · outbound
Generative OOD-regularized Model-based Policy Optimization Mitigating distribution shift in model-based offline RL via shifts-aware reward learning, 2024
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d8b49539-b33a-4b06-ad1e-650a06a351cb · outbound
Generative OOD-regularized Model-based Policy Optimization Supported value regular- ization for offline reinforcement learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 278b5253-d003-4191-ad0f-1e872e034493 · outbound
Generative OOD-regularized Model-based Policy Optimization Normalizing flows for inter- ventional density estimation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dc2076db-7f1a-44f6-a670-a27b753b92ac · outbound
Generative OOD-regularized Model-based Policy Optimization Model-based offline reinforcement learning with lower expectile q-learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1838f28c-ecc1-4ec2-b313-2b9ab54d917e · outbound
Generative OOD-regularized Model-based Policy Optimization Scalable diffusion models with transformers
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b0eb3366-c4a0-4426-8eed-e6f0d08e6f09 · outbound
Generative OOD-regularized Model-based Policy Optimization Continuous state-space models for optimal sepsis treatment: A deep reinforcement learning approach
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 01996bc2-7e99-43ad-91f5-baade0ce607b · outbound
Generative OOD-regularized Model-based Policy Optimization Variational inference with normalizing flows
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f8c537de-953a-4ad5-9b81-7cf661a90b40 · outbound
Generative OOD-regularized Model-based Policy Optimization High- resolution image synthesis with latent diffusion models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a0edb121-1f3f-4cfe-b5c4-6b04e730150b · outbound
Generative OOD-regularized Model-based Policy Optimization Kauffmann, Robert A
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c9558b8b-3788-4163-ae5c-510a90d68115 · outbound
Generative OOD-regularized Model-based Policy Optimization Denoising diffusion implicit models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 79ee89e0-5f82-4ed6-bf67-f95fce87043e · outbound
Generative OOD-regularized Model-based Policy Optimization Model-Bellman inconsistency for model-based offline reinforcement learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation db7f256e-795c-47f0-9337-c553b3206a24 · outbound
Generative OOD-regularized Model-based Policy Optimization Machine learning and imaging informatics in oncology.Oncology, 98(6):344–362, 2020
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 30f492f5-f168-425f-98c9-20c7bda894ca · outbound
Generative OOD-regularized Model-based Policy Optimization Guardian-regularized safe offline reinforcement learning for smart weaning of mechanical circulatory devices, 2025
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3e14cfd4-70a9-452f-bc51-638be52211bd · outbound
Generative OOD-regularized Model-based Policy Optimization Cardiogenic shock.Journal of the American Heart Association, 8(8):e011991, 2019
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 862ca1ce-18b3-47ad-8dd7-b1ccb990482e · outbound
Generative OOD-regularized Model-based Policy Optimization Diffusion policies as an expressive policy class for offline reinforcement learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 63ae57bb-51c3-4a55-9b85-2078c624801c · outbound
Generative OOD-regularized Model-based Policy Optimization Supported policy optimization for offline reinforcement learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9cb0c757-26ed-42f2-8b43-f52cf478019d · outbound
Generative OOD-regularized Model-based Policy Optimization Behavior regularized offline reinforcement learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 370656b6-dd92-43e4-b3af-3251b060d561 · outbound
Generative OOD-regularized Model-based Policy Optimization Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 55939b37-5cc0-4ea5-babb-00f5eedb4fdd · outbound
Generative OOD-regularized Model-based Policy Optimization Offline guarded safe reinforcement learning for medical treatment optimization strategies
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1719358a-f680-4cff-a2dc-f916963b39a2 · outbound
Generative OOD-regularized Model-based Policy Optimization Zou, Sergey Levine, Chelsea Finn, and Tengyu Ma
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e7dd0404-017f-4725-b0b3-0ed8cbc57aea · outbound
Generative OOD-regularized Model-based Policy Optimization Deep structured energy based models for anomaly detection
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9f27471e-6c22-443b-9015-3e998df676bf · outbound
Generative OOD-regularized Model-based Policy Optimization Constrained policy optimization with explicit behavior density for offline reinforcement learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 524f01bc-e6fe-4e20-afa5-c28fba37b466 · outbound
Generative OOD-regularized Model-based Policy Optimization OOD datasets
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9230c225-c12b-4c11-9565-d51d1a1fb477 · outbound
Generative OOD-regularized Model-based Policy Optimization Guidelines: • The answer [N/A] means that the paper does not involve crowdsourcing nor research with human subjects
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.