Pith. sign in

Paper Citation Record · LEDGER

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood

As of 14 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.08417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08417 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:22:27.653919Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact5
  • verified fuzzy28
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2727ca4-f5f9-430c-bae1-d53ed89ade7b · outbound

This paper cites write newline.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.652759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.652759Z digest=sha256:d6091b0fa067242e83917617686cb018c05cbd51aad61914d33b0df29150f2b4

Observation 04f87bd3-0ea1-41e8-bf93-3fa6826f2fb7 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified q-ensemble.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty-based offline reinforcement learning with diversified q-ensemble

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.687083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.687083Z digest=sha256:e27a82c3bc6ab33a2f9b0e944ba258497df7cc114ec37ec2f966c5de2dd8cd0b

Observation b62d4df3-f0fa-4a6b-9e96-2a1417d4dbb4 · outbound

This paper cites Near-optimal regret bounds for reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Near-optimal regret bounds for reinforcement learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.626467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.746503Z digest=sha256:ac3853fcddf620373724c4f78b40602af2f744e03315c4235617cbdbbd20f825

Observation 0012dfda-f904-4179-9f74-4956d0a95440 · outbound

This paper cites Manifold topology divergence: a framework for comparing data manifolds.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Manifold topology divergence: a framework for comparing data manifolds

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.611573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.776646Z digest=sha256:268d55ff95077be4550fd07b1946b3947225c50c3ac2d05f33b6ac769994092b

Observation 77245729-e20d-448a-8673-a1ab89bc83d5 · outbound

This paper cites Laplacian eigenmaps and spectral techniques for embedding and clustering.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Laplacian eigenmaps and spectral techniques for embedding and clustering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.848799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.848799Z digest=sha256:556ecf7a19fc0bc37e6e7934284d7ff6b112f4b5bab2dc6cb7dbd4d9ea41fb8b

Observation ad686d95-9820-434d-8f85-954a04ad2cb1 · outbound

This paper cites On the Inductive Bias of Neural Tangent Kernels.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the Inductive Bias of Neural Tangent Kernels

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.919649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.919649Z digest=sha256:68922ef69d62ec2a8321eef7f27454bb0aa254765f5ffe5c819d47bf592cc48e

Observation 77b6dbda-ac72-4049-bbe7-f8caf6c48683 · outbound

This paper cites Flows for simultaneous manifold learning and density estimation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Flows for simultaneous manifold learning and density estimation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.586125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.989185Z digest=sha256:3b0f1dfae8bc508c9a369845cc63d2521c46caa084ba223cd0d2c88e7120927a

Observation 36370ff4-c1f4-4016-bb26-f32426d32543 · outbound

This paper cites OpenAI Gym.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood OpenAI Gym

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.040414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.040414Z digest=sha256:ccd0a2ad2556c6909f49b15b1361c34c7320c34e7b33bdde27ef95c8a21a17bc

Observation edfe8984-2ac4-41d4-864d-aa29aaf70d34 · outbound

This paper cites Bail: Best-action imitation learning for batch deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Bail: Best-action imitation learning for batch deep reinforcement learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.097958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.097958Z digest=sha256:3fc0d0d540f6148dcc8ef5ddc79ece02315f5a0d5b188de75e5da9ae41502276

Observation c13f2ac5-941f-4bbc-ae58-d7ee4c981374 · outbound

This paper cites Diffusion maps.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Diffusion maps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.561407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.181646Z digest=sha256:dff1be095618bf5e476f9fa040039c5f2cc853096fbf3b1f76f6f76769a73c6f

Observation 46b1f972-6798-496f-a2c9-70fe477b9e2a · outbound

This paper cites Pink noise is all you need: Colored noise exploration in deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Pink noise is all you need: Colored noise exploration in deep reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.547034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.256430Z digest=sha256:02de75dea4bcc92374cd2f08acf11891854783eef79e5e7b76e66ffd09b1242d

Observation e57d3fa9-29c9-4044-89e8-69d03c6d1aef · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2021.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood D4rl: Datasets for deep data-driven reinforcement learning, 2021

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.311145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.311145Z digest=sha256:c9852f546ab97f9b50aea80ac029ec9f59d712f4357c9ed5fc8095dcb2ff1652

Observation 7ae79ce4-115e-4677-9086-b1f6e9c8af6d · outbound

This paper cites A minimalist approach to offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A minimalist approach to offline reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.522372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.363317Z digest=sha256:ddcf597c9497063159759d026ed36cf94ce628e4b669732c5ac07e3b86a03237

Observation 423edc72-3bbc-4e63-a4cb-c285062e1a18 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Off-policy deep reinforcement learning without exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.507852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.417323Z digest=sha256:2496a0bc07551b879c96ac4c6ecdc36717c1bc9cb6585c18ad6994973d776a4e

Observation 01b87718-6770-47eb-be7a-7a426b0c7700 · outbound

This paper cites Learning rankings via convex hull separation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning rankings via convex hull separation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.493119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.490677Z digest=sha256:73db78853ddb369d9fdbd230c7cc214171904865a4d0df82b8745ffd7863abfd

Observation 14dd046c-95d9-4c78-9362-9a3b7e400315 · outbound

This paper cites Extreme Q-Learning: MaxEnt RL without Entropy.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Extreme Q-Learning: MaxEnt RL without Entropy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.560434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.560434Z digest=sha256:02ab05f65c7aca9a354c3e013bcafbc73c12dc9b6b14b8afdeed6f8e41dd5ae4

Observation 0762145d-4da6-4aea-8d80-90ebb6106a89 · outbound

This paper cites Improving Offline RL by Blending Heuristics.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving Offline RL by Blending Heuristics

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:28.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.609399Z digest=sha256:53ad0ecc148f22ae63b4e035c5e50dc32b0ea1d87e0dbec39446674607c89966

Observation 5c018bd8-872b-4025-8e05-9254dd4f0ef2 · outbound

This paper cites Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.478149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.668798Z digest=sha256:2c584dbfe93430c2712c104caf108715d8f965101ce6545793a42a6187d3d97f

Observation 4e505ac3-d422-4752-b19e-2ae4a178802b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.748337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.748337Z digest=sha256:ad2de08f37d45ba55bbf3965da75421f21d86106ebf8a552018aabfb879a6548

Observation 087f99af-5649-49de-bd47-2c2e0f950ff8 · outbound

This paper cites Random projections for manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Random projections for manifold learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.451972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.810006Z digest=sha256:c54a9107f0e74330946adfb3e922512219c2bc386288efc18e33f62eb9b24faf

Observation f174c80b-b129-4129-8507-7b5ac451a2dd · outbound

This paper cites Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.436324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.868522Z digest=sha256:553dd98d72ba35110912112883e7d9d5e3bd840cce94d59ccedc66db2dd4e8a7

Observation 7483e25c-d457-4a7d-85ad-9f0c26e589c7 · outbound

This paper cites Mild policy evaluation for offline actor--critic.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mild policy evaluation for offline actor--critic

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.421916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.941158Z digest=sha256:4a6425644db7927b5b13deb9b977a62a4c1fa0a005ac932d959212a50b65d81f

Observation bb277ef5-c440-4243-8801-ae2ca4e47a57 · outbound

This paper cites Neural Tangent Kernel: Convergence and Generalization in Neural Networks.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Neural Tangent Kernel: Convergence and Generalization in Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.087828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.087828Z digest=sha256:a02950735a08e27c62cc6b4e7fe0f199ec96e79fb3daabee00389ec49a72c0b8

Observation 0971bad8-7bec-440e-91f0-f6a9c4586f28 · outbound

This paper cites A convex hull-based data selection method for data driven models.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A convex hull-based data selection method for data driven models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.407761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.188248Z digest=sha256:1cd9ef483b31ff531225a055c0e9d88b32c18bb4d7f2adf84536a6839aa709f3

Observation b435e2c8-7c51-4615-a19b-45bfc93b67ce · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.300203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.300203Z digest=sha256:404725ae8f1f46195a48e4eca5f42ded1f73fcae4c782ca27ae9c5ee2f39465b

Observation b4ac79f4-d8d5-4a13-8266-8afc8aabe59c · outbound

This paper cites Offline reinforcement learning with fisher divergence critic regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline reinforcement learning with fisher divergence critic regularization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.392898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.361489Z digest=sha256:2c40c88115ca4bc6ed66aed3fabf799b4d21234d42b4462aa8c0b3b60e752a49

Observation 2796467c-3a3c-4708-a2c5-91d74746154e · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline Reinforcement Learning with Implicit Q-Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.460008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.460008Z digest=sha256:4ac44494d52d18971939a6ef182ae7a81e1448ed950d97d22850318ea4b0fdd2

Observation 950a4f7b-b3a5-4673-bf86-e4a380591bf4 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.377984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.549357Z digest=sha256:070b36cb2e6b4c81143a5e2800a1bff514c523a898816a9d1ee67103a9e6554f

Observation 7acef59d-6aa9-4981-9cec-af33d8400982 · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Conservative q-learning for offline reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.669895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.669895Z digest=sha256:f0e16c56210f9d107c56fecb9f75721e33f805bc790b38dd51d25e36c377b377

Observation 8b551c13-6061-4f88-90c0-b6a1620cb95d · outbound

This paper cites Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.992371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.764096Z digest=sha256:a125aed48b75c4983d08e4e88c7b26d1b67b9c6cb5652eec9effb1ecb3b7c1e4

Observation 7a30ff04-8ad2-4ad5-a2df-3330feda9fd5 · outbound

This paper cites When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.971229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.844016Z digest=sha256:2925da48fc89353c08d29b98bbf3bb2a91ef3f8daaf389c1624cd88b92d8b49e

Observation 258495b4-b0ff-4cb4-99e9-f73d6b6800dc · outbound

This paper cites Mildly conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mildly conservative q-learning for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.352784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.916094Z digest=sha256:f143497f5d494d8a3d31db3a3c6a9b0ade69ae2458804daafd221ac9e83e2c68

Observation b067009a-5214-4801-85fc-419e7b097027 · outbound

This paper cites SEABO: A Simple Search-Based Method for Offline Imitation Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood SEABO: A Simple Search-Based Method for Offline Imitation Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.045813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.045813Z digest=sha256:af1829cda32bb1816680c99722c3465786ef43290457606abcc6422a69e68a4a

Observation 1196c970-65f4-4ac2-8829-4bdcd0bce33a · outbound

This paper cites On the role of general function approximation in offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the role of general function approximation in offline reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.337075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.115252Z digest=sha256:6e1ee7885ddb143f6ef601f0c2bbebbfe87dc1c1e486542f2ebd14806641e5cd

Observation 3d6c7838-3c2c-4f35-874e-6992398433af · outbound

This paper cites Machine learning algorithm based on convex hull analysis.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Machine learning algorithm based on convex hull analysis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.322758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.203111Z digest=sha256:de7b6a6bad6bef91279ed9e7e35ca5977b709dbf60082db95a8dbfca4d583617

Observation 44f94f89-08f2-41ca-be92-788441ea4c08 · outbound

This paper cites Why is Posterior Sampling Better than Optimism for Reinforcement Learning?.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why is Posterior Sampling Better than Optimism for Reinforcement Learning?

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.932706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.273819Z digest=sha256:fe9da3ed285c0fd7af342ebc064cc27142d21f86d4cc07e55e6afaac173c4d3d

Observation e3c0c889-bdd8-436e-ba64-193c5f90571d · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.320316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.320316Z digest=sha256:349ae75df51b7ef0cc6070cf1ba2dfbdbd4d3cae7ceb2580ac753c2f536b1c30

Observation 4fbc39b4-afb8-4167-9a6f-cd1fdce17504 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.393335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.393335Z digest=sha256:b5e9508bbf371afee000d2f78325daf66034fd642fc0e4c78a806e552d19a4b9

Observation 87e71b41-e966-446e-9f88-910d9bf76560 · outbound

This paper cites Policy regularization with dataset constraint for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Policy regularization with dataset constraint for offline reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.308226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.477113Z digest=sha256:30d46c1208506b9a5b5b368c7920873537917c4b640b8a64303d663d7a005216

Observation 1e868df6-a8bb-44bf-9b57-e2a19a62e438 · outbound

This paper cites Nonlinear dimensionality reduction by locally linear embedding.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Nonlinear dimensionality reduction by locally linear embedding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.293800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.530029Z digest=sha256:fc9c6dcbf125a26af64f2e089b1764aa56a6d0ffd65140f0d524d519ad6987fd

Observation e2562a35-92bd-4490-8447-166d9e165625 · outbound

This paper cites A dataset perspective on offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A dataset perspective on offline reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.279695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.572723Z digest=sha256:b97f9a24a1bb86f09496c014db243191ae3f773ec318d45882864c2cf09c72e7

Observation 5c6daea8-bc9b-4774-9770-09b969caaafe · outbound

This paper cites Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.632916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.632916Z digest=sha256:1d908527958554f8e15449e913bca1dc6eb34441ff8fce2d894bcd00e8784ac1

Observation be1e0c20-24d3-4c46-9125-6725309df06d · outbound

This paper cites Sutton and Andrew G.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Sutton and Andrew G

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.731486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.731486Z digest=sha256:e52d86c07d8e8474d7560ae373f4bb842318c1f7541b38a9d6778858e6de856f

Observation bdf784a3-85fc-4a18-a1b8-10cce6db4427 · outbound

This paper cites A global geometric framework for nonlinear dimensionality reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A global geometric framework for nonlinear dimensionality reduction

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.254341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.777449Z digest=sha256:f0d5fe883a67e6636847eaf5bfe40450a69d6a9740cf3e548ca8a9ab673671ad

Observation c3550023-30a5-4f75-860e-b65b43c43a7e · outbound

This paper cites Mujoco: A physics engine for model-based control.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mujoco: A physics engine for model-based control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.818469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.818469Z digest=sha256:d6703e7d64ad1d7e9f296c79c26fa3f9be831e3542f5024f99955170428c1ce8

Observation 97666142-1f76-4ba5-8747-d477965e14a5 · outbound

This paper cites Adaptive manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adaptive manifold learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.237822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.909702Z digest=sha256:9d36498f4b8263c5e30d5583bfa4e9d08947c8988839e552d5fd5b3713db156b

Observation c32c712c-f7b1-4f43-826c-ec7805ae0790 · outbound

This paper cites Improving generalization in reinforcement learning with mixture regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving generalization in reinforcement learning with mixture regularization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.223654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.975185Z digest=sha256:d2f58d04a00a4ef0e844d7b7503cb8fb8099a79a254ff08baa76ca335dd1e05c

Observation 6b0d68aa-5d91-45a0-872b-69df93598980 · outbound

This paper cites Exponentially weighted imitation learning for batched historical data.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Exponentially weighted imitation learning for batched historical data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.993358Z digest=sha256:0704daabab5ef60a195e7c1af2bf7edd4a451a4f8422dd0ce048db31ccaa552d

Observation 77c722db-48ec-49f6-b92f-dfe28046357c · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Behavior Regularized Offline Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.055708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.055708Z digest=sha256:7dbd3f57e1005d6c7c7296837665d12fa1a9278631a0318d57433acc5a898cf0

Observation fbb913a5-e59d-418c-a15d-31470aad2381 · outbound

This paper cites Zeta hull pursuits: Learning nonconvex data hulls.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Zeta hull pursuits: Learning nonconvex data hulls

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.194510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.159787Z digest=sha256:dc655809331cc7946421b501b2ed60b5bb8d84dc59628cd9f621e45cf405316e

Observation 45f17633-9104-4ffd-9565-f2fa0d64e7c0 · outbound

This paper cites Uncertainty svm active learning algorithm based on convex hull and sample distance.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty svm active learning algorithm based on convex hull and sample distance

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.179977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.217496Z digest=sha256:c9f6aa412896bea1c73e3c9e6153862591a566323b3781cfe8de623d92eba1dc

Observation a73e2a5f-1b28-4238-8d50-5f4e3eb823c2 · outbound

This paper cites Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.299725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.299725Z digest=sha256:adcb682cdc0ec2fb8da6084896773b57e4ceb502eaf9ed5f0649c0a899f011d3

Observation 29ef1edb-2f09-4290-8ace-a5f42f6e7ef8 · outbound

This paper cites Rorl: Robust offline reinforcement learning via conservative smoothing.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Rorl: Robust offline reinforcement learning via conservative smoothing

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.402534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.402534Z digest=sha256:820db18a31b71d3215d55cc5862855b5ba7e092efec7f688f216a30f7570bd6d

Observation a42cc906-eef6-49be-9a4f-edf5155a047b · outbound

This paper cites Towards Robust Offline Reinforcement Learning under Diverse Data Corruption.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Towards Robust Offline Reinforcement Learning under Diverse Data Corruption

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.698977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.451131Z digest=sha256:a97626f39b06aee41fb94b059be51e5eebf875a9b3409340adff3bae9d89ca07

Observation a0d3b574-9bee-46ee-8dc8-2945e3a3ccb4 · outbound

This paper cites In-sample actor critic for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood In-sample actor critic for offline reinforcement learning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.154600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.547421Z digest=sha256:2d20676138e130e5b68b503519d312909e834d100e03d5c3a11f753f970b91cc

Observation 2572a492-724e-4fa6-94af-d2a55b0b8b92 · outbound

This paper cites @esa (Ref.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood @esa (Ref

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.644839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.644839Z digest=sha256:6284edeb3b2ab3bf3083a258324195dc85ed2a5b094d79b0a0fafe49d4f79b01

Observation 772415cc-6c2a-46be-9dac-6ee889b10779 · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.649510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.649510Z digest=sha256:a1ba4070b9fd9d9a3b011884a2f122063e87f1566c3cda8bd5537cf88fe52243

Observation 9b3481cc-3210-418a-a671-83b889d1099b · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.653919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.653919Z digest=sha256:3d6bf7899de50c76cc1b524e4f7a6f9e48d9ce5e6cd86bb22d0dd8af94486a4c

Pith citing papers

No inbound Pith citation observations are available.