Pith. sign in

Paper Citation Record · LEDGER

Variational OOD State Correction for Offline Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.00503.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.00503 v3

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:46:08.077256Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact2
  • verified fuzzy16
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ef48b7a-14cd-479f-ab2e-d0d3145028b2 · outbound

This paper cites GPT-4 Technical Report.

Variational OOD State Correction for Offline Reinforcement Learning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.965220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.965220Z digest=sha256:9a0f602173694498741cb359dfe0eac571af7292d38359e873f571e6561b0e00

Observation 99096310-b63b-46a7-a373-48b63db62b94 · outbound

This paper cites Learning markov state abstractions for deep reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Learning markov state abstractions for deep reinforcement learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.400482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.969768Z digest=sha256:5f2b6f14868c7d4d1f0e5ffced51ac046d1ba61f16cb42983bc27bd533e2bfb6

Observation 31367436-eaa5-4c53-b449-1d95c54811b9 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified q-ensemble.

Variational OOD State Correction for Offline Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified q-ensemble

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.391492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.973100Z digest=sha256:b42789ac25267bfcfc446fdd2ef67b9221f293d6bf9f638c7dc00130ce1e81de

Observation dd567200-1253-42c1-9a0c-313ef739d610 · outbound

This paper cites Pessimistic bootstrapping for uncertainty-driven offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Pessimistic bootstrapping for uncertainty-driven offline reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.382105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.976836Z digest=sha256:e8f9b1055645d4a8faa27d9c8ad00a4f51befd9a9f246a6a3e6774dacb07439b

Observation bc4da37b-efec-4f24-8847-de23489ccacf · outbound

This paper cites Label-noise robust logistic regression and its applications.

Variational OOD State Correction for Offline Reinforcement Learning Label-noise robust logistic regression and its applications

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.372433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.980276Z digest=sha256:6fbd5e65c839157a46fa6008c67c261d41538a9356fa27c7a989a7aa425651f4

Observation 644e8d65-2e9b-424a-9f37-13d8059d9c18 · outbound

This paper cites Importance Weighted Autoencoders.

Variational OOD State Correction for Offline Reinforcement Learning Importance Weighted Autoencoders

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.983646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.983646Z digest=sha256:3089c1b7a1638b89631480c48a897a0161a0fefc113e246b0753bb823ff922bb

Observation c74fccd9-e3bd-4b09-a8ff-dffe82f11d8f · outbound

This paper cites Tutorial on Variational Autoencoders.

Variational OOD State Correction for Offline Reinforcement Learning Tutorial on Variational Autoencoders

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.987957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.987957Z digest=sha256:2c382596f659f3c4deeec4c9db1f82c949f04fddbbb89f87a01f36076122d8e3

Observation 77265d7d-19f6-4a15-a776-b59715b6960a · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Variational OOD State Correction for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.991451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.991451Z digest=sha256:347af89817ffe8c1f83531ff68089452de896b1782f0888b62cd72973b57e54f

Observation e17f339a-c7d0-449b-8255-1e612dc0fda8 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Variational OOD State Correction for Offline Reinforcement Learning Off-policy deep reinforcement learning without exploration

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.362697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.994764Z digest=sha256:0ae3d908e9a7f275d80feb86da4e99247c013e7ade18bb7dd41c8265ffc0a858

Observation aee7d1f4-b677-4eb2-892e-53ce6e4efd57 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning A comprehensive survey on safe reinforcement learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.998083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.998083Z digest=sha256:6a80be857f4d957d43c8aaa599dbc42d70c05fd38972f84106a7562e8713ec55

Observation 05514fed-c6af-47bf-bfc3-e0f6155478a2 · outbound

This paper cites Estimation of non-normalized statistical models by score matching.

Variational OOD State Correction for Offline Reinforcement Learning Estimation of non-normalized statistical models by score matching

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.001245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.001245Z digest=sha256:3261835fa1ac4b4bc7323b519c1e5043b493f21963b2c55509be4c1e5707705f

Observation fc48874b-b5b7-4575-aec1-3c9f4866fadb · outbound

This paper cites Planning with Diffusion for Flexible Behavior Synthesis.

Variational OOD State Correction for Offline Reinforcement Learning Planning with Diffusion for Flexible Behavior Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.008559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.008559Z digest=sha256:73b1320c54980e488d8c099e2fc561ba533909ec01d9c5e0e9a8ae04c4e8917e

Observation d3e7bd93-c663-4fc8-8a28-9a4b8ed1cabb · outbound

This paper cites Recovering from out-of-sample states via inverse dynamics in offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Recovering from out-of-sample states via inverse dynamics in offline reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.342922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.012362Z digest=sha256:83b084d4b88233c28bd3a0f9496369a22ba72d44c4c4e814802ace87783247ef

Observation 7a0d9678-f707-4c44-ba55-1a10b9a9fca4 · outbound

This paper cites an unresolved cited work.

Variational OOD State Correction for Offline Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:46:08.332759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.015471Z digest=sha256:3a38943363d4dd2f9067f783238f2405577f39f43802937e0462a137e4d294c0

Observation cc2807a4-8eb1-4fe9-9019-3ca614338c95 · outbound

This paper cites Scalable deep reinforcement learning for vision-based robotic manipulation.

Variational OOD State Correction for Offline Reinforcement Learning Scalable deep reinforcement learning for vision-based robotic manipulation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.323263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.018719Z digest=sha256:a242ec69f5a0214a10e1e46e4e4a8dced790f755bbea7006cb4e18566789fb10

Observation 8fd905cf-c544-4936-8ada-8068da5e5e3c · outbound

This paper cites Lyapunov density models: Constraining distribution shift in learning-based control.

Variational OOD State Correction for Offline Reinforcement Learning Lyapunov density models: Constraining distribution shift in learning-based control

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.313028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.021758Z digest=sha256:d89d26da82e9b3d602183eecb7b7040343e1a715985f76c95e026210c2550682

Observation 23404f0f-5c81-45da-8c2f-ba52a23a7863 · outbound

This paper cites Kingma and Max Welling.

Variational OOD State Correction for Offline Reinforcement Learning Kingma and Max Welling

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.024790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.024790Z digest=sha256:905ed67a6d8ad0ff9f9451cde30512183ac52962eeab1a8d826f5332ef679d82

Observation 85b7faeb-8dbe-47b5-8987-fdfc68a02a9c · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Conservative q-learning for offline reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.297346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.027953Z digest=sha256:23a49bc86ff644abbaf00056ad5e74e69382bd28c88c58854dbed14ca756e523

Observation 397c65bb-9786-48bd-899d-038cbc8dda7c · outbound

This paper cites Batch reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Batch reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.030931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.030931Z digest=sha256:96afe79f72b602c47c0e5d552b45a81af3eaa49e2f40fb83ed377296a673e255

Observation c9ada6e4-f5ef-48db-9be9-6964eee85dc7 · outbound

This paper cites Supported value regularization for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Supported value regularization for offline reinforcement learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.280727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.034368Z digest=sha256:b67201125992111382e00837ae8a40438217d5e8738e5538a2d1601ebc262d33

Observation 0aae39b5-9e6a-454d-94cf-706f7adb7df4 · outbound

This paper cites Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression.

Variational OOD State Correction for Offline Reinforcement Learning Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:46:08.144011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.037469Z digest=sha256:5d5a8afd1ca9f8432acca7da4c35f1b00fe3195a54e0f3df9c99ab739d1aafd0

Observation 2e559b6e-9885-4097-9294-ffad781c7a30 · outbound

This paper cites Human-level control through deep reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Human-level control through deep reinforcement learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.040956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.040956Z digest=sha256:7465ba60e7081971fa258e342c102a4049876b2d15cca60f42fa087b57232cd3

Observation 092274a0-d39b-42ab-a467-a2bc360b27fc · outbound

This paper cites Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.265226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.043948Z digest=sha256:3e01a12b9931fc9f62f8f864a52e105133168773e51d86f504ef4ab5d77c1c9f

Observation e98c70ba-ad82-4ed3-b7b0-01ac023280e3 · outbound

This paper cites Mastering the game of go without human knowledge.

Variational OOD State Correction for Offline Reinforcement Learning Mastering the game of go without human knowledge

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.046845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.046845Z digest=sha256:114216f2a536e1e356ec536bca05029f18a138d1a3beba48ec142c27a0418c2b

Observation 34534397-3fd8-466f-b144-d77d96c76e07 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Variational OOD State Correction for Offline Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.049612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.049612Z digest=sha256:92be2bde14cff8f015cae0796342266ef269ab25c81cc8a9c39d19e44cd6ef58

Observation 0d1d9d1b-7911-43d5-bc75-7ddcca1d3607 · outbound

This paper cites Robust distance metric learning in the presence of label noise.

Variational OOD State Correction for Offline Reinforcement Learning Robust distance metric learning in the presence of label noise

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.249511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.052689Z digest=sha256:d2cd940361521dee2893cabe18900e0f175559113c288197c3c6b00cfec75058

Observation 83c57b72-b982-4f67-bfa2-35f8d884eb2f · outbound

This paper cites an unresolved cited work.

Variational OOD State Correction for Offline Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:46:08.238963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.055612Z digest=sha256:6a43dac98b007a98386d2d4ca76cf72efd8185c6bd139a11bc2d5dc050d729fc

Observation 43b9d0b8-28e4-46c7-aa6e-fba8f1034228 · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Variational OOD State Correction for Offline Reinforcement Learning Behavior Regularized Offline Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.058685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.058685Z digest=sha256:bc6da0ddf0588b7f14298567b679d4ea5b025511482dc19f97dc28ca2dda278e

Observation 13475ed8-4071-40a5-bbb8-dbfd4fe2687d · outbound

This paper cites Supported policy optimization for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Supported policy optimization for offline reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.229077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.061838Z digest=sha256:d2438f7291ff75786e7993fa2ec1ddcd01b720a8df2859f2aaef6767da9bda27

Observation 1999d70c-3f12-4a5c-b282-7d01cecce048 · outbound

This paper cites RORL: Robust Offline Reinforcement Learning via Conservative Smoothing.

Variational OOD State Correction for Offline Reinforcement Learning RORL: Robust Offline Reinforcement Learning via Conservative Smoothing

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:46:08.113098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.064871Z digest=sha256:d72918337c9ba54f5031ba74b9e0a28139f63a79cbe1bb6d09c212c4134ce6f6

Observation a879cd6a-38f7-4ccf-a223-a454c0989d89 · outbound

This paper cites An implicit trust region approach to behavior regularized offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning An implicit trust region approach to behavior regularized offline reinforcement learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.218677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.068085Z digest=sha256:0ba2575ab1be9fd40d6e11c908ed5c0bd9fd5db59a8e971a923a5083327b622a

Observation 4ea89976-a352-4dfa-b40c-aba3d976a6fe · outbound

This paper cites State deviation correction for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning State deviation correction for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.208816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.071301Z digest=sha256:75d7295bb1ec9527058df6a37792197011be65cf7fc1b7747f27240b92c1ace0

Observation 9840f7b5-c279-4ecf-aa8d-79d284f69b90 · outbound

This paper cites Constrained policy optimization with explicit behavior density for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Constrained policy optimization with explicit behavior density for offline reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.199344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.074192Z digest=sha256:19db41d1d7cc1c4308d5f58091f69c3a2e47c9d41d89be19dc4f9b05d255c4ec

Observation 85c0bc90-d052-4f8d-b25f-85ed27f8bca7 · outbound

This paper cites write newline.

Variational OOD State Correction for Offline Reinforcement Learning write newline

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.077256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.077256Z digest=sha256:fb7bf001f18e83efe011e3c37197e3d57325f8bd534fcaca8f58354382a60fb9

Pith citing papers

No inbound Pith citation observations are available.