Pith. sign in

Paper Citation Record · LEDGER

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling

As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2502.06491.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06491 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:57.334192Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 026cb979-7686-4b92-8880-2ee95fe78810 · outbound

This paper cites an unresolved cited work.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:57.743081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.203760Z digest=sha256:7fc0d0c5ac6ff197197503025c0eb7afd4d0afaaa49d3baca36d99a8fa0327a9

Observation 55d23ecb-21c4-4d8d-989e-ef35a4696425 · outbound

This paper cites De- cision transformer: Reinforcement learning via sequence modeling.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling De- cision transformer: Reinforcement learning via sequence modeling

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.705723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.216414Z digest=sha256:9391fc1d4e96aa6af640a66c97ba893637ada36e25f7223af5fd798a02bcce9f

Observation 4269d9f1-b62e-46d1-bdca-f7fb7b3a5dd8 · outbound

This paper cites UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:57.223190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:22:57.223190Z digest=sha256:bcd1497eb1b7ed11d6a3ff486592eb233d102e000896b1765a51cbb048125433

Observation 70be4e63-6abf-45d9-ab32-ecd9e84a69c1 · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:57.226625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:22:57.226625Z digest=sha256:2f2eb84cb45086fc98f85160ea498ddb541c4c1a533110e2ed5451aa8b2ac085

Observation 29e74946-14b7-4ce1-95f6-9cd394fd2d09 · outbound

This paper cites Bridging the data gap between training and inference for unsupervised neural machine translation.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Bridging the data gap between training and inference for unsupervised neural machine translation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.670597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.235914Z digest=sha256:47d481d1bccf204032b40024d98fd13792a31cc53c7106a9e0f87725f1ab327f

Observation 18f60e09-4984-470e-b489-ca6c1d27ce53 · outbound

This paper cites Offline reinforcement learning as one big se- quence modeling problem.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline reinforcement learning as one big se- quence modeling problem

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.645403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.245241Z digest=sha256:7ad7c0f4873af1f13328513dd9177afce8cf8c4ee271198fec8032f593b5fad6

Observation 6280b86b-f4ff-4442-9386-8508074da6a1 · outbound

This paper cites CTRL: A Conditional Transformer Language Model for Controllable Generation.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling CTRL: A Conditional Transformer Language Model for Controllable Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:57.248534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:22:57.248534Z digest=sha256:05e9c43123e593d2aa934f75bab547e4293ac5e5881e323f434d71b690890aae

Observation 7f3a3ed9-5e91-4949-9333-ac83f829dc08 · outbound

This paper cites Morel: Model-based offline reinforcement learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Morel: Model-based offline reinforcement learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.636829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.252753Z digest=sha256:400292396668ff34c7f92cf06cf388d3a00a4230c11c732934b1489400c03163

Observation 5c4afab4-e449-4a73-9b09-458d5cbbe853 · outbound

This paper cites Offline reinforcement learning with im- plicit q-learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline reinforcement learning with im- plicit q-learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.627969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.256445Z digest=sha256:ce8745b83a11cc868954db72ef4630f253424115f639f921cae207802bf068dc

Observation 41586ed2-9973-4f06-a537-f562519308fd · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Conservative q-learning for offline reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.619426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.259885Z digest=sha256:04cb086ca839bf951cdb9daf880cc4c6af841b2fc424941c130cf330517b82d0

Observation 9a951ff5-7ddd-471d-a61b-333113962602 · outbound

This paper cites Multi-game decision transformers.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Multi-game decision transformers

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.611076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.263329Z digest=sha256:d9af2fd61118ca4d3b308caea6de8fc0b09909c97572fa284fe590ea44a4c4f0

Observation 717ee278-1cab-4ba7-832c-ac210c4d99b9 · outbound

This paper cites Distribution- conditioned adversarial variational autoencoder for valid instrumental variable generation.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Distribution- conditioned adversarial variational autoencoder for valid instrumental variable generation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.601910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.266857Z digest=sha256:0c395817d203bb2f092692e63e9615b426d6c74da312225080db7ec9d90e7eed

Observation 92cd01e8-920c-40c4-be8c-f1d1247e33d2 · outbound

This paper cites Diffstitch: Boosting of- fline reinforcement learning with diffusion-based trajec- tory stitching.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Diffstitch: Boosting of- fline reinforcement learning with diffusion-based trajec- tory stitching

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.592871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.270618Z digest=sha256:0c3a870f059454eda39fb835145010e46cc42512442adec1882022f2e329873a

Observation c7f7a13b-1aec-4776-9d28-dcef19612457 · outbound

This paper cites Ball, Yee Whye Teh, and Jack Parker-Holder.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Ball, Yee Whye Teh, and Jack Parker-Holder

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.583625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.273848Z digest=sha256:20783662aa4889d447826f344c2dbda7b7214af965c07d6374c687a656ca01e7

Observation fe2073c9-2ccb-4673-8f77-654dfba6d71c · outbound

This paper cites Luis, Alessandro G.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Luis, Alessandro G

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.575212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.277371Z digest=sha256:b2d2068daea89d17aaaa952de1b6451cf88060172440a8080fae35c2ea56cc90

Observation 1950a005-210e-4aeb-96cf-7037685de849 · outbound

This paper cites Double check your state before trusting it: Confidence- aware bidirectional offline model-based imagination.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Double check your state before trusting it: Confidence- aware bidirectional offline model-based imagination

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.566186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.280821Z digest=sha256:5f0e85bd7e38ae6283df831e0b514a056bfba183a2a81e8a62219740234071e4

Observation 61e7ecd6-f73c-4b00-9986-25798328309e · outbound

This paper cites Reining generalization in offline re- inforcement learning via representation distinction.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Reining generalization in offline re- inforcement learning via representation distinction

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.556304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.284177Z digest=sha256:79e2bbbf490ac41106e6f3e0457bfec2dd61b2710261a21d8e9d47b0c47aab0f

Observation 8c61a2e2-8a3a-498f-9eed-d19572ee22c4 · outbound

This paper cites Offline Imitation Learning with Model-based Reverse Augmentation.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline Imitation Learning with Model-based Reverse Augmentation

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:22:57.369429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.287578Z digest=sha256:80857d4bca876cb8d91769acdb43e906cadeaa01dd2da5b09cd82b3f71bde247

Observation 01e72fb7-2537-4b69-8fa2-00159d6e4547 · outbound

This paper cites Model-bellman in- consistency for model-based offline reinforcement learn- ing.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Model-bellman in- consistency for model-based offline reinforcement learn- ing

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.547316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.291356Z digest=sha256:c53ee699b484a66aecff8fc692acdcfbd64e72ba3dca394a02af32c325d061ba

Observation c2d81fe2-a9c1-4f6d-8151-5914df624f47 · outbound

This paper cites an unresolved cited work.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:57.538430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.294834Z digest=sha256:dc6ad605d7fb0caf9175bd0dda30ff1648d8390cb5ac8b4a31e8941b80e6a640

Observation 15d7fb2a-a021-40a5-b218-a061e4b2056b · outbound

This paper cites Self-correcting models for model-based reinforcement learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Self-correcting models for model-based reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.529437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.298237Z digest=sha256:dd80da359bf89f2b3b81c060311617b65942726cba921a7617bb7f4df9df7c24

Observation b78e3e17-89ba-4638-b1eb-edf7f8c15c02 · outbound

This paper cites Visualizing data using t-sne.Journal of machine learning research, 9(11),.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Visualizing data using t-sne.Journal of machine learning research, 9(11),

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:57.304653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:22:57.304653Z digest=sha256:de00ee78517073672add50d999ef453e34efe4044151de17e435dbe7f5c34056

Observation 56d541da-ebea-4604-b467-bd85d01fe907 · outbound

This paper cites Offline reinforcement learning with reverse model-based imagination.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Offline reinforcement learning with reverse model-based imagination

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.495528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.310939Z digest=sha256:709b035e38029f0735ec29393e8b5836ba9c683ef9ea8c9120a133dabcd5c8f5

Observation c4cb93a7-0fb8-41f1-81f3-58e809ab790d · outbound

This paper cites Q-learning decision transformer: Leveraging dynamic programming for conditional se- quence modelling in offline RL.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Q-learning decision transformer: Leveraging dynamic programming for conditional se- quence modelling in offline RL

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.484638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.313868Z digest=sha256:4c2e83f0cf6aef9aed722ae6f73258b27af2056826aeb87e76cc85bde721e126

Observation d188fd78-b098-42f4-8718-2de07b8da280 · outbound

This paper cites Pareto policy pool for model-based offline reinforcement learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Pareto policy pool for model-based offline reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.475757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.316920Z digest=sha256:f92ad05cf59e583b64be0f2bdec14e0450472c622d2bae93bf59764ebe056e43

Observation cde9e3bd-eef7-43ed-8821-d20a5a4b7949 · outbound

This paper cites Zou, Sergey Levine, Chelsea Finn, and Tengyu Ma.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Zou, Sergey Levine, Chelsea Finn, and Tengyu Ma

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.464744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.320106Z digest=sha256:997143b8fdd5e93ce3c404f9dbe030da8a29a97be6b98571cd0add55399ac707

Observation 0a3006d4-6613-4f3f-91c4-5dfd042a72a8 · outbound

This paper cites Combo: Conservative offline model-based policy opti- mization.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Combo: Conservative offline model-based policy opti- mization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.453045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.322865Z digest=sha256:3d464f50df19862357cd46809787089e25627659cbc8bf2937d668610349c362

Observation 4ec02951-1bfa-467f-8cb0-8740c0533f86 · outbound

This paper cites Model-based offline planning with trajectory prun- ing.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Model-based offline planning with trajectory prun- ing

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.442149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.325731Z digest=sha256:a5a3fce7f81275a55b1579625887b8c79e122b3d85590791fa46407120e6cc1a

Observation c8406b5e-fd84-4b11-90de-47f9c3e0f528 · outbound

This paper cites Uncertainty-driven trajectory truncation for data augmen- tation in offline reinforcement learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Uncertainty-driven trajectory truncation for data augmen- tation in offline reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.430685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.328642Z digest=sha256:daaee11c6c88ad439f5d682f1748fa67042640d2bc244bd1661ebcc1f79293f6

Observation c1b29786-5bdb-4317-9ae6-90c9dca0a147 · outbound

This paper cites Conditional vari- ational autoencoder for sign language translation with cross-modal alignment.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Conditional vari- ational autoencoder for sign language translation with cross-modal alignment

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.419360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.331555Z digest=sha256:9dafd510626e175684c483c1bbb55b343e2a233192ebf79b1573903655f3c8a0

Observation 13d9d2d0-9298-4b86-92c8-54bdfe8e946b · outbound

This paper cites Is model ensemble necessary? model- based RL via a single model with lipschitz regularized value function.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Is model ensemble necessary? model- based RL via a single model with lipschitz regularized value function

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.407879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.334192Z digest=sha256:25850f2ece4b8d625e7c420a311bf2de8aa2d3e17e82172e41d90b0dbf6c2f22

Observation 556ad3d2-8dce-4ba1-b319-0d167b3e30f2 · outbound

This paper cites Jamieson.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Jamieson

Reference 2008

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.505853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.307781Z digest=sha256:40ef24e252caa1a889648be5737cd4bbb2effae56c2c3f7bc1727c0c1987e75c

Observation a880dac2-8a39-4295-ad3c-93c285a6fd89 · outbound

This paper cites When to trust your model: Model-based policy optimization.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling When to trust your model: Model-based policy optimization

Reference 2010

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.653888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.242035Z digest=sha256:1bee67fdfc4c7bc3d11e13bc0b8de1e281fdbdd74fb2bb9ef6a5e50c11a06fd2

Observation fc4a1364-2801-4845-9c84-56e87ae577b8 · outbound

This paper cites Behavioral cloning from observation.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Behavioral cloning from observation

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.520050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.301527Z digest=sha256:f78348847cbfa776ce085510587ab5519528cf801297701a08b6a29c3e035fab

Observation 8c6bcd7a-3cab-481b-92bc-ee2992815790 · outbound

This paper cites Waypoint trans- former: Reinforcement learning via supervised learning with intermediate targets.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Waypoint trans- former: Reinforcement learning via supervised learning with intermediate targets

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.733466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.207356Z digest=sha256:4ec302039c9b0bc36fc85da154d6bae972cd73b793bd493ed6324486dd370729

Observation 166a5bd9-fec1-4769-959d-2c9967232283 · outbound

This paper cites ACT: empowering decision transformer with dynamic program- ming via advantage conditioning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling ACT: empowering decision transformer with dynamic program- ming via advantage conditioning

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.679007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.232925Z digest=sha256:3d10bdff1ab1a91f96aba65a13452b65f2d0dab8743d551b7d5fb1b8e00e5aa6

Observation 96854808-e76e-48ee-a68b-f5274a2c64e7 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Off-policy deep reinforcement learning without exploration

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.687089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.230025Z digest=sha256:907c40ff3ef1ab4fd02ae232cf553c590ce1574405153f863be9fcc04a3fd7e3

Observation 31eb5e0d-8161-47fc-a12a-57f08cbf02b9 · outbound

This paper cites Lapo: Latent-variable advantage-weighted policy optimization for offline rein- forcement learning.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Lapo: Latent-variable advantage-weighted policy optimization for offline rein- forcement learning

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.695616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.219766Z digest=sha256:ca9e6213ff086898ceb88b990aea25b39fc2d3ad563689eaccb77898dbbdc9dc

Observation f3934870-e60d-4ca6-bc41-691e8e65e092 · outbound

This paper cites an unresolved cited work.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Unresolved cited work

Reference 2022

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:57.662108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.239136Z digest=sha256:c9d65be3b62c0df79a9cc1fbd52ed91d09f5827ec621ebf5926e3edde344fdec

Observation 56ec4d35-2a57-4f8c-8796-f22eeffb4fc8 · outbound

This paper cites Data-efficient task generalization via probabilistic model-based meta reinforcement learn- ing.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Data-efficient task generalization via probabilistic model-based meta reinforcement learn- ing

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.724076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.210404Z digest=sha256:302a1e6837688686daa7da7d0572f2ac22bc68676f4f535ee88198a65bdf1420

Observation 4b3ba3be-fe11-49f6-9117-8c0278e0b5c8 · outbound

This paper cites Blanchet, Miao Lu, Tong Zhang, and Han Zhong.

Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling Blanchet, Miao Lu, Tong Zhang, and Han Zhong

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:57.715280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T15:22:57.213411Z digest=sha256:b92e8fad10680533d63476068a1a269a2f40b447e8b175ef5fe1c36bfc179755

Pith citing papers

No inbound Pith citation observations are available.