Pith. sign in

Paper Citation Record · LEDGER

Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2209.14548.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.14548 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:28:41.720110Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:17:36.926223Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bc403606-e921-4844-87c0-1257f34c67ad · inbound

IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies cites this paper.

IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:48:36.454455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T13:48:36.369334Z digest=sha256:c8e4a2c34fa7bdadb2c88d207e80eadc5360d9cd95722a2a273d423ff5e803c1

Observation 9f6973f3-5d8a-48a4-bcfe-65f9796f460d · inbound

Diffusion Policy Policy Optimization cites this paper.

Diffusion Policy Policy Optimization Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:48:14.849919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T08:48:14.776754Z digest=sha256:cd6034936366ae057e708300ed2daaef579ff9e103d1cc7f41855f765da3efda

Observation d50c7894-b470-476d-bfe2-3e82e52ff6cc · inbound

Efficient Online Reinforcement Learning for Diffusion Policy cites this paper.

Efficient Online Reinforcement Learning for Diffusion Policy Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T19:28:41.720110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:28:41.720110Z digest=sha256:56ec35e0d2d8d0dfc415965a29402e2309b865ae2a1cb9960c9c26e7f25aaad9

Observation 1e34ce56-c691-4043-9a86-7da0dad1c6e4 · inbound

Skill Expansion and Composition in Parameter Space cites this paper.

Skill Expansion and Composition in Parameter Space Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:09.250585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:09.250585Z digest=sha256:39d1abdf94941ebc2e18bb8649e9cf7abb743cdd6c28ba5290e58b5fb2dc74a8

Observation ad21c487-7e76-4c69-aab3-0e6db50e2a36 · inbound

Decision Flow Policy Optimization cites this paper.

Decision Flow Policy Optimization Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:20:36.514447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:20:36.514447Z digest=sha256:59dba0c40b096eec340dcdfb151b6c5f79e70689f5b2df7c725b8732de3f7add

Observation 3543fcc2-b22f-4209-9493-448ac32654fc · inbound

Steering Your Diffusion Policy with Latent Space Reinforcement Learning cites this paper.

Steering Your Diffusion Policy with Latent Space Reinforcement Learning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:55:46.454030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T21:55:46.183007Z digest=sha256:b409d19e16df130732d0c0c09d9be5055de96f0268636fcca87304388ecd1ca9

Observation dede0fb2-7726-4036-bc49-c3a9da87a142 · inbound

Offline Reinforcement Learning with Penalized Action Noise Injection cites this paper.

Offline Reinforcement Learning with Penalized Action Noise Injection Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:38:31.961205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:38:31.961205Z digest=sha256:a3b7c475f0cec6673cbd21bcfe8e60baf91bbb05d27607d00b28f1eddac55d85

Observation a2ac85bb-43bc-4c6c-9919-df83144e5dd0 · inbound

EXPO: Stable Reinforcement Learning with Expressive Policies cites this paper.

EXPO: Stable Reinforcement Learning with Expressive Policies Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:05.295957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T05:09:02.111308Z digest=sha256:e6da5a27ef950964defd35bf79c1fc87145c6e6583da5c1923d45ad1fa0c819c

Observation e342b956-78bb-454c-bfc1-7531e970994b · inbound

Value Flows cites this paper.

Value Flows Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T11:01:27.541506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:01:27.541506Z digest=sha256:4ad2bbe1667fae5396adf374f0b44ddeea076f09e803441dd0cc55114442a9bc

Observation 6c79076d-6a69-431a-9878-fd404a2d9fee · inbound

From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning cites this paper.

From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:40:34.113297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T00:36:32.681470Z digest=sha256:1fdd39155b3bf01dce85c441b343fadd44b5878a95a97d1581136eed285b7634

Observation 9a521830-861d-4c6c-98d1-22073219b0e6 · inbound

From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning cites this paper.

From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:04:19.102993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T19:02:38.240098Z digest=sha256:506fe001733668cf6fa0e057ff9eafb1a0e184caf6ad108ddbbcb90b671fc123

Observation d718eb6a-e545-4792-a915-594ba3139dfa · inbound

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning cites this paper.

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T23:46:32.301737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:46:32.301737Z digest=sha256:8fcbb13a883124ddd33ad9ba8ba82bbb31d0ed3cf36a8b958e19cff9a7bc6371

Observation e0843021-e8b2-4877-a260-e2615928d6f4 · inbound

FASTER: Value-Guided Sampling for Fast RL cites this paper.

FASTER: Value-Guided Sampling for Fast RL Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:48:27.243217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T02:47:36.475845Z digest=sha256:5e21c8527a676b32cd4619542c4722abcd21aadedcb3aba639844827156f68ba

Observation da6727d5-62b2-4592-ad0b-486eaa77abc1 · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:56:00.756160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T16:25:25.739019Z digest=sha256:43bb1b71d4bda47f6ba06d3b2e4cc3bcc9c305583a29b1d66edf25b11c9542e7

Observation 74428c6e-5b63-440d-8899-8bf6e43b269f · inbound

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning cites this paper.

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:41:50.895414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T02:17:25.783688Z digest=sha256:40d812ad4c285aa2606aa7577ea78b7ff6881fa0acfe71709c295debf2c3b01c

Observation 7301ba23-4bce-4553-931d-ceadc608acd2 · inbound

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies cites this paper.

Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T17:22:24.711336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T17:19:36.722892Z digest=sha256:1dc4b8565fd55626bf2e198d585ea42a33c2ae6546db44b87d64569d5cb76eac

Observation 55acafa4-4dab-43da-a28a-b9d88668067f · inbound

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios cites this paper.

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:08.571497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:47:10.062975Z digest=sha256:f392e036f7208d7373f1af0b22a8c3e2126cbfe8dcbdf540a5a6d64da54b9240

Observation 56a45925-a7f3-4bee-8473-16d4b3df509f · inbound

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning cites this paper.

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:17:36.927634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T14:05:01.073951Z digest=sha256:94c49048d862b015d43c8a9a169bb216da6e5d42fa3fc4a2ad9c5061181d3aae

Observation d5b9b3bc-3de7-4331-ac35-a086475da3a2 · inbound

A Single Diffusion-Policy Controller for Multi-Task Block Pushing with Zero-Shot Sim-to-Real Transfer cites this paper.

A Single Diffusion-Policy Controller for Multi-Task Block Pushing with Zero-Shot Sim-to-Real Transfer Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T08:30:10.584105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:30:10.584105Z digest=sha256:597a54c42c83ce9008daf857a88bed48400127ec3388dadc9402d766aad4fdb4