Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:57.334192Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2502.06491.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:57.334192Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 026cb979-7686-4b92-8880-2ee95fe78810 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 55d23ecb-21c4-4d8d-989e-ef35a4696425 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling De- cision transformer: Reinforcement learning via sequence modeling
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4269d9f1-b62e-46d1-bdca-f7fb7b3a5dd8 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70be4e63-6abf-45d9-ab32-ecd9e84a69c1 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29e74946-14b7-4ce1-95f6-9cd394fd2d09 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Bridging the data gap between training and inference for unsupervised neural machine translation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 18f60e09-4984-470e-b489-ca6c1d27ce53 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline reinforcement learning as one big se- quence modeling problem
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6280b86b-f4ff-4442-9386-8508074da6a1 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling CTRL: A Conditional Transformer Language Model for Controllable Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f3a3ed9-5e91-4949-9333-ac83f829dc08 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Morel: Model-based offline reinforcement learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5c4afab4-e449-4a73-9b09-458d5cbbe853 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline reinforcement learning with im- plicit q-learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 41586ed2-9973-4f06-a537-f562519308fd · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Conservative q-learning for offline reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9a951ff5-7ddd-471d-a61b-333113962602 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Multi-game decision transformers
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 717ee278-1cab-4ba7-832c-ac210c4d99b9 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Distribution- conditioned adversarial variational autoencoder for valid instrumental variable generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 92cd01e8-920c-40c4-be8c-f1d1247e33d2 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Diffstitch: Boosting of- fline reinforcement learning with diffusion-based trajec- tory stitching
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c7f7a13b-1aec-4776-9d28-dcef19612457 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Ball, Yee Whye Teh, and Jack Parker-Holder
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fe2073c9-2ccb-4673-8f77-654dfba6d71c · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Luis, Alessandro G
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1950a005-210e-4aeb-96cf-7037685de849 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Double check your state before trusting it: Confidence- aware bidirectional offline model-based imagination
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 61e7ecd6-f73c-4b00-9986-25798328309e · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Reining generalization in offline re- inforcement learning via representation distinction
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8c61a2e2-8a3a-498f-9eed-d19572ee22c4 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline Imitation Learning with Model-based Reverse Augmentation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 01e72fb7-2537-4b69-8fa2-00159d6e4547 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Model-bellman in- consistency for model-based offline reinforcement learn- ing
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c2d81fe2-a9c1-4f6d-8151-5914df624f47 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 15d7fb2a-a021-40a5-b218-a061e4b2056b · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Self-correcting models for model-based reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b78e3e17-89ba-4638-b1eb-edf7f8c15c02 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Visualizing data using t-sne.Journal of machine learning research, 9(11),
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56d541da-ebea-4604-b467-bd85d01fe907 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline reinforcement learning with reverse model-based imagination
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c4cb93a7-0fb8-41f1-81f3-58e809ab790d · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Q-learning decision transformer: Leveraging dynamic programming for conditional se- quence modelling in offline RL
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d188fd78-b098-42f4-8718-2de07b8da280 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Pareto policy pool for model-based offline reinforcement learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cde9e3bd-eef7-43ed-8821-d20a5a4b7949 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Zou, Sergey Levine, Chelsea Finn, and Tengyu Ma
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0a3006d4-6613-4f3f-91c4-5dfd042a72a8 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Combo: Conservative offline model-based policy opti- mization
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4ec02951-1bfa-467f-8cb0-8740c0533f86 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Model-based offline planning with trajectory prun- ing
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c8406b5e-fd84-4b11-90de-47f9c3e0f528 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Uncertainty-driven trajectory truncation for data augmen- tation in offline reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c1b29786-5bdb-4317-9ae6-90c9dca0a147 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Conditional vari- ational autoencoder for sign language translation with cross-modal alignment
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 13d9d2d0-9298-4b86-92c8-54bdfe8e946b · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Is model ensemble necessary? model- based RL via a single model with lipschitz regularized value function
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 556ad3d2-8dce-4ba1-b319-0d167b3e30f2 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Jamieson
Reference 2008
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a880dac2-8a39-4295-ad3c-93c285a6fd89 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling When to trust your model: Model-based policy optimization
Reference 2010
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fc4a1364-2801-4845-9c84-56e87ae577b8 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Behavioral cloning from observation
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8c6bcd7a-3cab-481b-92bc-ee2992815790 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Waypoint trans- former: Reinforcement learning via supervised learning with intermediate targets
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 166a5bd9-fec1-4769-959d-2c9967232283 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling ACT: empowering decision transformer with dynamic program- ming via advantage conditioning
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 96854808-e76e-48ee-a68b-f5274a2c64e7 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Off-policy deep reinforcement learning without exploration
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 31eb5e0d-8161-47fc-a12a-57f08cbf02b9 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Lapo: Latent-variable advantage-weighted policy optimization for offline rein- forcement learning
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f3934870-e60d-4ca6-bc41-691e8e65e092 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 56ec4d35-2a57-4f8c-8796-f22eeffb4fc8 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Data-efficient task generalization via probabilistic model-based meta reinforcement learn- ing
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4b3ba3be-fe11-49f6-9117-8c0278e0b5c8 · outbound
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Blanchet, Miao Lu, Tong Zhang, and Han Zhong
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.