Pith. sign in

Paper Citation Record · LEDGER

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood

As of 14 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.08417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08417 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:22:27.653919Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact5
  • verified fuzzy28
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2727ca4-f5f9-430c-bae1-d53ed89ade7b · outbound

This paper cites write newline.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.652759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.652759Z digest=sha256:d6091b0fa067242e83917617686cb018c05cbd51aad61914d33b0df29150f2b4

Observation 04f87bd3-0ea1-41e8-bf93-3fa6826f2fb7 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified q-ensemble.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty-based offline reinforcement learning with diversified q-ensemble

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.687083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.687083Z digest=sha256:e27a82c3bc6ab33a2f9b0e944ba258497df7cc114ec37ec2f966c5de2dd8cd0b

Observation b62d4df3-f0fa-4a6b-9e96-2a1417d4dbb4 · outbound

This paper cites Near-optimal regret bounds for reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Near-optimal regret bounds for reinforcement learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.626467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.746503Z digest=sha256:cb9086e916abdfea4e94eef95afd5324ea572d3a62eb9e8225b453110c341ca4

Observation 0012dfda-f904-4179-9f74-4956d0a95440 · outbound

This paper cites Manifold topology divergence: a framework for comparing data manifolds.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Manifold topology divergence: a framework for comparing data manifolds

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.611573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.776646Z digest=sha256:74c498970d972231fc4f344726ca8bb08b430fac3f05eeccd15e259eb738de66

Observation 77245729-e20d-448a-8673-a1ab89bc83d5 · outbound

This paper cites Laplacian eigenmaps and spectral techniques for embedding and clustering.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Laplacian eigenmaps and spectral techniques for embedding and clustering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.848799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.848799Z digest=sha256:556ecf7a19fc0bc37e6e7934284d7ff6b112f4b5bab2dc6cb7dbd4d9ea41fb8b

Observation ad686d95-9820-434d-8f85-954a04ad2cb1 · outbound

This paper cites On the Inductive Bias of Neural Tangent Kernels.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the Inductive Bias of Neural Tangent Kernels

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.919649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.919649Z digest=sha256:a06e43788a891b661ac4d3b0fbba585c27f71b44888d0e9607a1d6c990c093ce

Observation 77b6dbda-ac72-4049-bbe7-f8caf6c48683 · outbound

This paper cites Flows for simultaneous manifold learning and density estimation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Flows for simultaneous manifold learning and density estimation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.586125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.989185Z digest=sha256:c09b28c6ec5dcbbf6166c820152f466e17f24f2810f473f2d580c8c0f65f6970

Observation 36370ff4-c1f4-4016-bb26-f32426d32543 · outbound

This paper cites OpenAI Gym.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood OpenAI Gym

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.040414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.040414Z digest=sha256:ccd0a2ad2556c6909f49b15b1361c34c7320c34e7b33bdde27ef95c8a21a17bc

Observation edfe8984-2ac4-41d4-864d-aa29aaf70d34 · outbound

This paper cites Bail: Best-action imitation learning for batch deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Bail: Best-action imitation learning for batch deep reinforcement learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.097958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.097958Z digest=sha256:3fc0d0d540f6148dcc8ef5ddc79ece02315f5a0d5b188de75e5da9ae41502276

Observation c13f2ac5-941f-4bbc-ae58-d7ee4c981374 · outbound

This paper cites Diffusion maps.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Diffusion maps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.561407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.181646Z digest=sha256:7173db4715225dfea6e07bf13ba8f150b2d5d3c185d37515a54310209f1fa6c4

Observation 46b1f972-6798-496f-a2c9-70fe477b9e2a · outbound

This paper cites Pink noise is all you need: Colored noise exploration in deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Pink noise is all you need: Colored noise exploration in deep reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.547034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.256430Z digest=sha256:791eb4177ca84a44cd66a3f0cc817b84fbbb18fe6150b88ccf85f356e3b30c0d

Observation e57d3fa9-29c9-4044-89e8-69d03c6d1aef · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2021.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood D4rl: Datasets for deep data-driven reinforcement learning, 2021

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.311145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.311145Z digest=sha256:c9852f546ab97f9b50aea80ac029ec9f59d712f4357c9ed5fc8095dcb2ff1652

Observation 7ae79ce4-115e-4677-9086-b1f6e9c8af6d · outbound

This paper cites A minimalist approach to offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A minimalist approach to offline reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.522372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.363317Z digest=sha256:c90b89a50fcdefe0f685c0dabecf57a720a1c61d731ea033dd41bd6c6a323c0b

Observation 423edc72-3bbc-4e63-a4cb-c285062e1a18 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Off-policy deep reinforcement learning without exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.507852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.417323Z digest=sha256:cdc94b5bc4e5e73d57e22f1538f5edf1f549641ce05194027de4ccc274a93d22

Observation 01b87718-6770-47eb-be7a-7a426b0c7700 · outbound

This paper cites Learning rankings via convex hull separation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning rankings via convex hull separation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.493119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.490677Z digest=sha256:624421de49e26483aecc4beb91624345344e94d1a773d9260c021ada079a3b81

Observation 14dd046c-95d9-4c78-9362-9a3b7e400315 · outbound

This paper cites Extreme Q-Learning: MaxEnt RL without Entropy.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Extreme Q-Learning: MaxEnt RL without Entropy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.560434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.560434Z digest=sha256:02ab05f65c7aca9a354c3e013bcafbc73c12dc9b6b14b8afdeed6f8e41dd5ae4

Observation 0762145d-4da6-4aea-8d80-90ebb6106a89 · outbound

This paper cites Improving Offline RL by Blending Heuristics.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving Offline RL by Blending Heuristics

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:28.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.609399Z digest=sha256:0bdfa2480e74a790085142aa09b278446dbc1b96c5d4e460895ef8d3df969fc0

Observation 5c018bd8-872b-4025-8e05-9254dd4f0ef2 · outbound

This paper cites Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.478149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.668798Z digest=sha256:9e470ec94b7cde6a335275c2c74b02e59a3f19360020be18c7bd0277080e6373

Observation 4e505ac3-d422-4752-b19e-2ae4a178802b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.748337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.748337Z digest=sha256:ad2de08f37d45ba55bbf3965da75421f21d86106ebf8a552018aabfb879a6548

Observation 087f99af-5649-49de-bd47-2c2e0f950ff8 · outbound

This paper cites Random projections for manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Random projections for manifold learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.451972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.810006Z digest=sha256:92f4a8bae115620feb5d0515f0041d4d8a462aee6c9e6b152ef42ef7d6f9cd39

Observation f174c80b-b129-4129-8507-7b5ac451a2dd · outbound

This paper cites Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.436324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.868522Z digest=sha256:26c7e328d903a5792fd3d825977dc2dc46c514ca38aa9411d2223d1cac662037

Observation 7483e25c-d457-4a7d-85ad-9f0c26e589c7 · outbound

This paper cites Mild policy evaluation for offline actor--critic.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mild policy evaluation for offline actor--critic

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.421916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.941158Z digest=sha256:c6e2d12d6b27fa681537e9424bdaf42a5ce727a44c43e1a603d597c38d2da95e

Observation bb277ef5-c440-4243-8801-ae2ca4e47a57 · outbound

This paper cites Neural Tangent Kernel: Convergence and Generalization in Neural Networks.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Neural Tangent Kernel: Convergence and Generalization in Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.087828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.087828Z digest=sha256:a02950735a08e27c62cc6b4e7fe0f199ec96e79fb3daabee00389ec49a72c0b8

Observation 0971bad8-7bec-440e-91f0-f6a9c4586f28 · outbound

This paper cites A convex hull-based data selection method for data driven models.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A convex hull-based data selection method for data driven models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.407761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.188248Z digest=sha256:d763cc1924f7ad5f2ac4fe47780347249c8935ba403f99615ee7a6ba4a7e54e3

Observation b435e2c8-7c51-4615-a19b-45bfc93b67ce · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.300203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.300203Z digest=sha256:4043f101b96ab5f04e055fba09dc9b425f23003e430a64d3a5a4ea915eaa3a72

Observation b4ac79f4-d8d5-4a13-8266-8afc8aabe59c · outbound

This paper cites Offline reinforcement learning with fisher divergence critic regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline reinforcement learning with fisher divergence critic regularization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.392898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.361489Z digest=sha256:f6bcfd34b8d675da4c7173317963383e371b7ecd47ee4cb8c2ae65430a946c34

Observation 2796467c-3a3c-4708-a2c5-91d74746154e · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline Reinforcement Learning with Implicit Q-Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.460008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.460008Z digest=sha256:4ac44494d52d18971939a6ef182ae7a81e1448ed950d97d22850318ea4b0fdd2

Observation 950a4f7b-b3a5-4673-bf86-e4a380591bf4 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.377984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.549357Z digest=sha256:f5418724cadf767d3532994c85c050e5fd3fc8af3aa35195d259bb0c2a0db458

Observation 7acef59d-6aa9-4981-9cec-af33d8400982 · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Conservative q-learning for offline reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.669895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.669895Z digest=sha256:f0e16c56210f9d107c56fecb9f75721e33f805bc790b38dd51d25e36c377b377

Observation 8b551c13-6061-4f88-90c0-b6a1620cb95d · outbound

This paper cites Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.992371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.764096Z digest=sha256:98912519863aae99f61c6d21ed3e1e6ffdb75a8958f3f13ec3935678de9ba45f

Observation 7a30ff04-8ad2-4ad5-a2df-3330feda9fd5 · outbound

This paper cites When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.971229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.844016Z digest=sha256:dffb9cf66c985a6ea4246ff2ba2a726d293ffdaba72746b03fd434f2086eff44

Observation 258495b4-b0ff-4cb4-99e9-f73d6b6800dc · outbound

This paper cites Mildly conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mildly conservative q-learning for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.352784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.916094Z digest=sha256:afea3753565cab9108c1a2a76df47d81a0b778ecd0f68ee65f167716579a603c

Observation b067009a-5214-4801-85fc-419e7b097027 · outbound

This paper cites SEABO: A Simple Search-Based Method for Offline Imitation Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood SEABO: A Simple Search-Based Method for Offline Imitation Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.045813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.045813Z digest=sha256:af1829cda32bb1816680c99722c3465786ef43290457606abcc6422a69e68a4a

Observation 1196c970-65f4-4ac2-8829-4bdcd0bce33a · outbound

This paper cites On the role of general function approximation in offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the role of general function approximation in offline reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.337075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.115252Z digest=sha256:2c7a08c25ee58791c85cd4e0116e254eb5ba786c17c9085f749a1e6f33f7d584

Observation 3d6c7838-3c2c-4f35-874e-6992398433af · outbound

This paper cites Machine learning algorithm based on convex hull analysis.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Machine learning algorithm based on convex hull analysis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.322758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.203111Z digest=sha256:a232a00b5d3af75059f68928500ae4a24feba3a20624f815b4526547eea99b6e

Observation 44f94f89-08f2-41ca-be92-788441ea4c08 · outbound

This paper cites Why is Posterior Sampling Better than Optimism for Reinforcement Learning?.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why is Posterior Sampling Better than Optimism for Reinforcement Learning?

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.932706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.273819Z digest=sha256:9b044fe28ebfbb3d0b5d80397fa8a9c2bf6420d08ab8ccb1cff3ba74f7ac29db

Observation e3c0c889-bdd8-436e-ba64-193c5f90571d · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.320316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.320316Z digest=sha256:349ae75df51b7ef0cc6070cf1ba2dfbdbd4d3cae7ceb2580ac753c2f536b1c30

Observation 4fbc39b4-afb8-4167-9a6f-cd1fdce17504 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.393335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.393335Z digest=sha256:b5e9508bbf371afee000d2f78325daf66034fd642fc0e4c78a806e552d19a4b9

Observation 87e71b41-e966-446e-9f88-910d9bf76560 · outbound

This paper cites Policy regularization with dataset constraint for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Policy regularization with dataset constraint for offline reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.308226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.477113Z digest=sha256:120fdbb52dce752469df82e345a46b5fdf22d0f857976a099c183dbff4b16941

Observation 1e868df6-a8bb-44bf-9b57-e2a19a62e438 · outbound

This paper cites Nonlinear dimensionality reduction by locally linear embedding.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Nonlinear dimensionality reduction by locally linear embedding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.293800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.530029Z digest=sha256:9a57806c09355fdfc5d56377da4828f88e2f148a73d3129d786b0f13116fc729

Observation e2562a35-92bd-4490-8447-166d9e165625 · outbound

This paper cites A dataset perspective on offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A dataset perspective on offline reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.279695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.572723Z digest=sha256:4fab5edc937f1e676d49ed6d296ac4e1415aa7fa4591d4fc1fa2e6fbe4538040

Observation 5c6daea8-bc9b-4774-9770-09b969caaafe · outbound

This paper cites Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.632916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.632916Z digest=sha256:c2582effdc895a3265079e54fdf1394f48089e61b820fe429b85d7846878e245

Observation be1e0c20-24d3-4c46-9125-6725309df06d · outbound

This paper cites Sutton and Andrew G.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Sutton and Andrew G

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.731486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.731486Z digest=sha256:e52d86c07d8e8474d7560ae373f4bb842318c1f7541b38a9d6778858e6de856f

Observation bdf784a3-85fc-4a18-a1b8-10cce6db4427 · outbound

This paper cites A global geometric framework for nonlinear dimensionality reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A global geometric framework for nonlinear dimensionality reduction

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.254341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.777449Z digest=sha256:af4729876a5b82cfea1485c8a60c4b044393825f7512fd25c8f867f6196cd46d

Observation c3550023-30a5-4f75-860e-b65b43c43a7e · outbound

This paper cites Mujoco: A physics engine for model-based control.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mujoco: A physics engine for model-based control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.818469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.818469Z digest=sha256:d6703e7d64ad1d7e9f296c79c26fa3f9be831e3542f5024f99955170428c1ce8

Observation 97666142-1f76-4ba5-8747-d477965e14a5 · outbound

This paper cites Adaptive manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adaptive manifold learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.237822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.909702Z digest=sha256:f602fb2f459c5f4967f46169462007b2b72e2e26b52f928140a04dca6cf3fe88

Observation c32c712c-f7b1-4f43-826c-ec7805ae0790 · outbound

This paper cites Improving generalization in reinforcement learning with mixture regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving generalization in reinforcement learning with mixture regularization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.223654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.975185Z digest=sha256:5e1fc216845dc52b76b3c268ea32469c7422451310f904b8c1b5e7471743a300

Observation 6b0d68aa-5d91-45a0-872b-69df93598980 · outbound

This paper cites Exponentially weighted imitation learning for batched historical data.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Exponentially weighted imitation learning for batched historical data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.993358Z digest=sha256:7f9449830f2b4ed49caf261622fafa69aa803cf7e75948b52f0cbc9b1aedae5b

Observation 77c722db-48ec-49f6-b92f-dfe28046357c · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Behavior Regularized Offline Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.055708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.055708Z digest=sha256:7dbd3f57e1005d6c7c7296837665d12fa1a9278631a0318d57433acc5a898cf0

Observation fbb913a5-e59d-418c-a15d-31470aad2381 · outbound

This paper cites Zeta hull pursuits: Learning nonconvex data hulls.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Zeta hull pursuits: Learning nonconvex data hulls

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.194510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.159787Z digest=sha256:60794030c4b069171a131e5e409bbcf1cda757f2dfec8c53570a03a0d2764818

Observation 45f17633-9104-4ffd-9565-f2fa0d64e7c0 · outbound

This paper cites Uncertainty svm active learning algorithm based on convex hull and sample distance.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty svm active learning algorithm based on convex hull and sample distance

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.179977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.217496Z digest=sha256:9940404dbd770417301b2b0f23dd5e6d6d0fc212c01464a65cebf334c1d9cbfd

Observation a73e2a5f-1b28-4238-8d50-5f4e3eb823c2 · outbound

This paper cites Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.299725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.299725Z digest=sha256:adcb682cdc0ec2fb8da6084896773b57e4ceb502eaf9ed5f0649c0a899f011d3

Observation 29ef1edb-2f09-4290-8ace-a5f42f6e7ef8 · outbound

This paper cites Rorl: Robust offline reinforcement learning via conservative smoothing.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Rorl: Robust offline reinforcement learning via conservative smoothing

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.402534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.402534Z digest=sha256:820db18a31b71d3215d55cc5862855b5ba7e092efec7f688f216a30f7570bd6d

Observation a42cc906-eef6-49be-9a4f-edf5155a047b · outbound

This paper cites Towards Robust Offline Reinforcement Learning under Diverse Data Corruption.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Towards Robust Offline Reinforcement Learning under Diverse Data Corruption

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.698977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.451131Z digest=sha256:0ea3b8d95c48b90e16e26e1a561ea5d40b545ae006ea37fb64a0df4157410ff4

Observation a0d3b574-9bee-46ee-8dc8-2945e3a3ccb4 · outbound

This paper cites In-sample actor critic for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood In-sample actor critic for offline reinforcement learning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.154600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.547421Z digest=sha256:b17a26658c35dbb3e39278a256cd08a0b411f109aeb1852a3e9ed9bdff3dcba7

Observation 2572a492-724e-4fa6-94af-d2a55b0b8b92 · outbound

This paper cites @esa (Ref.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood @esa (Ref

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.644839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.644839Z digest=sha256:6284edeb3b2ab3bf3083a258324195dc85ed2a5b094d79b0a0fafe49d4f79b01

Observation 772415cc-6c2a-46be-9dac-6ee889b10779 · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.649510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.649510Z digest=sha256:a1ba4070b9fd9d9a3b011884a2f122063e87f1566c3cda8bd5537cf88fe52243

Observation 9b3481cc-3210-418a-a671-83b889d1099b · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.653919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.653919Z digest=sha256:3d6bf7899de50c76cc1b524e4f7a6f9e48d9ce5e6cd86bb22d0dd8af94486a4c

Pith citing papers

No inbound Pith citation observations are available.