Pith. sign in

Paper Citation Record · LEDGER

Exploratory Diffusion Model for Unsupervised Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2502.07279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.07279 v2

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:21:18.474336Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact1
  • verified fuzzy28
  • unresolved40
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 717804df-2c64-4226-83eb-f028019cddf1 · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Deep reinforcement learning at the edge of the statistical precipice

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.552117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.114277Z digest=sha256:315045f4a8e81019468073551dbbb757cd5e976755d59b41cada6a2102d8bf9b

Observation 1186c289-b8cd-485a-8f62-13bd7c4d93da · outbound

This paper cites Tenenbaum, Tommi S.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Tenenbaum, Tommi S

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.120405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.120405Z digest=sha256:68b4c58623ffa98bde99775b8c0348182da4be7034ca790b142f59db8f5abb3b

Observation 52013a13-f6fb-4638-a35d-f2950327fed6 · outbound

This paper cites Diffusion for World Modeling: Visual Details Matter in Atari.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion for World Modeling: Visual Details Matter in Atari

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.125931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.125931Z digest=sha256:5e77198759f6e09611866f63fe56bdf876a1bd635503f2ca47314eb93020c507

Observation 1c04bd1b-5546-4893-bbf7-dd6bb3d3540b · outbound

This paper cites Random polytopes, convex bodies, and approximation.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Random polytopes, convex bodies, and approximation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.526532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.132825Z digest=sha256:4cb4971afd097bf8a3f248b7ac1d6dde04587ef1d907cd8202e33bf58a275ea3

Observation b5f258a2-ebde-445b-a142-3e292e93ebe1 · outbound

This paper cites Constrained Ensemble Exploration for Unsupervised Skill Discovery.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Constrained Ensemble Exploration for Unsupervised Skill Discovery

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-08T13:21:18.939754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.138225Z digest=sha256:5894cb8f70085da0bf130feac2b681005922cca7edd687d1e8122cd3a8e0cf52

Observation 8f7e1583-b62b-47d7-81c5-d76e8d9c4e00 · outbound

This paper cites Exploration by random network distillation.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Exploration by random network distillation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.144431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.144431Z digest=sha256:fe830348b2abad945656cba537c0bec4d81a6ab088fc26f36160f488d0ea8c94

Observation 9ce0c823-a0ac-4384-888a-ad82f96ec597 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Explore, discover and learn: Unsupervised discovery of state-covering skills

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.500380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.150828Z digest=sha256:41736dc094065f527097a092bbfe9af1a2d937a8f3f5006de85cb542bd04b1b3

Observation 4f41f3bf-524c-4d7d-b5ec-4ce13f26c927 · outbound

This paper cites DIME:Diffusion-Based Maximum Entropy Reinforcement Learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning DIME:Diffusion-Based Maximum Entropy Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.155791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.155791Z digest=sha256:3936713b16b84f5395f91c685f79d4f430d06c1b637c58cd981164082e1fb7d0

Observation 749b2cf6-799b-4bcd-ad89-0bd4e41b92a5 · outbound

This paper cites Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.160941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.160941Z digest=sha256:4a91dec99477fb9be37a2ac6c6164226bf023f63f8082815055cc4b2f9dc64c5

Observation 193876ed-7802-427e-94db-c299f5ddd8df · outbound

This paper cites Simple Hierarchical Planning with Diffusion.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Simple Hierarchical Planning with Diffusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.166804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.166804Z digest=sha256:29a86bd2347bfac40d29d881a968c620b0d951c5230ba816fd26d8fe844e9669

Observation 249abdf1-5bc1-4cb2-b12b-91895894bbe9 · outbound

This paper cites Offline reinforcement learning via high-fidelity generative behavior modeling.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Offline reinforcement learning via high-fidelity generative behavior modeling

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.484341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.172641Z digest=sha256:71c063ff123843b3f650f6a8d2fee6252f5cdc0063330d4602e64d72286889f8

Observation b00d0d5a-5bcb-4ec0-81b1-cf47f5dbac27 · outbound

This paper cites Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.177946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.177946Z digest=sha256:8b56e946090fb8193318d9a761bbd56b63a92a0023e084809579706e11fca65b

Observation 20988d13-2fcd-4b7a-9bad-507db77e9192 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion policy: Visuomotor policy learning via action diffusion

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.183415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.183415Z digest=sha256:71cb221ac3c8570178812e4a1b1a50b8d2e2314d070ca096cf7dc516b29bbefe

Observation b8dca37e-e625-40fc-98d4-dffeb92cb3b0 · outbound

This paper cites Diffusion Posterior Sampling for General Noisy Inverse Problems.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Posterior Sampling for General Noisy Inverse Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.188814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.188814Z digest=sha256:8a7c269226172813426b0d65fa53fe09bf5fe8a9a098350e8e83560eec883627

Observation f1b9aff1-ba8b-4410-81e3-dbf0334efd1a · outbound

This paper cites Diffusion models beat gans on image synthesis.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion models beat gans on image synthesis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.194341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.194341Z digest=sha256:9a0709214f381ab14a269e7fad68a0bd15c92bfbe8d6551b61337166a46aaf87

Observation 5bc01a3b-00f9-4a54-837c-210713bbf508 · outbound

This paper cites Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.199380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.199380Z digest=sha256:b500f85e85eb5d73e62bfbe2e3a3bbc8a5bb0d97064e1224f1280eff2876bba4

Observation cba315ae-fa8a-4bee-abeb-1f10bb41e019 · outbound

This paper cites Diversity is all you need: Learning skills without a reward function.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diversity is all you need: Learning skills without a reward function

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.447422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.204573Z digest=sha256:8ab2cafd85750e1783032550d9b5804be44ed16a0632ddf1c236ffe10dc7e955

Observation 54de6cbc-71f0-46b0-a13f-32ea678baa77 · outbound

This paper cites The information geometry of unsupervised reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning The information geometry of unsupervised reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.431559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.209501Z digest=sha256:70fd6f2339528ebe55957e47ac6475f203fd38b5587bd449ecbd07f0e29ae0ed

Observation e6ded4df-fa16-481d-ac4b-24d104174cae · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Reinforcement learning with deep energy-based policies

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.416165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.214553Z digest=sha256:a914ae63230b600aa33b670b4d532b3aae58e744d225a9081e2044bbdfbb172a

Observation 8b9f12ab-cea0-4ef4-984c-bbdd091494d1 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.400294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.219668Z digest=sha256:c8adc1099b136af50476401c91aa850af035fb64834e5f6ced656ceb1c92afd0

Observation 1803d2cc-0e49-4e91-8a24-4828acd0e416 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.224350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.224350Z digest=sha256:20658d7dd821d6a665f340062a71391b624ee6816c253474c277cf844ec7a0d9

Observation d90fbec2-cc75-4562-a0d7-519a7e761017 · outbound

This paper cites Diffusion model is an effective planner and data synthesizer for multi-task reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion model is an effective planner and data synthesizer for multi-task reinforcement learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.384186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.229805Z digest=sha256:c90702911cbe3e7be676f415e12b09319834f36ff7fa129068966461fb35053c

Observation 65d6ffed-8431-4bc1-a236-6b8da9905ac9 · outbound

This paper cites Denoising diffusion probabilistic models.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Denoising diffusion probabilistic models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.234847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.234847Z digest=sha256:a72f428f3ab80566537d8ff98782d5269cddf0f598266b8982213a0dcaf76764

Observation 09ba4fbd-d15c-448d-86d4-586560b38a47 · outbound

This paper cites Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.239564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.239564Z digest=sha256:82af8afb2e3109bf3e0e0c444d239fb4478cf77f7cd8e9d4c5b0a8d6ab3950e0

Observation 64ea5b7d-a4d1-46a6-b3d7-bc072f27085f · outbound

This paper cites Planning with diffusion for flexible behavior synthesis.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Planning with diffusion for flexible behavior synthesis

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.358510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.245569Z digest=sha256:a8c9c48d5cf8a76e55725b3c5331147dbeca3d4ade97cb46fc55af72b86dc435

Observation 0c81c733-6952-4b95-b4ef-2d8cb7a0c569 · outbound

This paper cites Efficient diffusion policies for offline reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Efficient diffusion policies for offline reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.250301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.250301Z digest=sha256:09a5ecfb2bf38af54431a947c58b1278c04498c17cd37b16d4b859766354001c

Observation 3204f4ed-570d-404d-a18a-3d0ca8e4e2a1 · outbound

This paper cites Unsupervised skill discovery with bottleneck option learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Unsupervised skill discovery with bottleneck option learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.333288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.255235Z digest=sha256:bf896c375d42cd359dc460ed799a9e89407c5b4c059f3ffef9811c4120ac87c1

Observation edc71b78-a83a-4795-9c54-6c4638cda4b5 · outbound

This paper cites Offline reinforcement learning with implicit q-learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Offline reinforcement learning with implicit q-learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.259982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.259982Z digest=sha256:ece0727c24382602bf5758e3dc95f0d5d69bf3d2e913402991a4cbb82107dd3b

Observation 73b8f274-5633-4882-8931-202c37755d65 · outbound

This paper cites Unsuper- vised reinforcement learning with contrastive intrinsic control.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Unsuper- vised reinforcement learning with contrastive intrinsic control

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.306391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.264755Z digest=sha256:9954959924e87a79f5835f7ebae2053f548bcd6eab8cedf5121b40036c6ad19c

Observation cf259c37-eb98-47a3-ad23-3cb9c6052bb5 · outbound

This paper cites Urlb: Unsupervised reinforcement learning benchmark.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Urlb: Unsupervised reinforcement learning benchmark

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.289871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.269581Z digest=sha256:5e5f87447350ac79779951a0110a412326c4fef7511ebfe50b6af83a39dc81be

Observation b4104a14-fb61-4278-a931-3d2289529bc0 · outbound

This paper cites Efficient Exploration via State Marginal Matching.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Efficient Exploration via State Marginal Matching

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.274465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.274465Z digest=sha256:c072e7f424983d8ad3cb0e3b1f2def42f2890845318993996dde55eb35ba4687

Observation 5c091b6d-b0e9-499a-b58b-8156c8a0d9ec · outbound

This paper cites Hierarchical diffusion for offline decision making.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Hierarchical diffusion for offline decision making

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.272069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.279537Z digest=sha256:916ae25cb8ac6176cfd20f1c1b762cdd653cde741b38bf58ed8162eb2c7ccd4e

Observation 5d47d839-47e1-4e91-8bf6-b317946c851b · outbound

This paper cites Learning Multimodal Behaviors from Scratch with Diffusion Policy Gradient.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Learning Multimodal Behaviors from Scratch with Diffusion Policy Gradient

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.284794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.284794Z digest=sha256:80da4abe4daddc796ebfa6dacfa138ae3700db02022592ef350b96cb0182c8b1

Observation 6c2cff4f-45cc-4e34-84b6-a9c51d4e44ad · outbound

This paper cites Adaptdiffuser: Diffusion models as adaptive self-evolving planners.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Adaptdiffuser: Diffusion models as adaptive self-evolving planners

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.256958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.290217Z digest=sha256:6b40b8546d0db5ecac72620ef7cf02ab1fba4590b086b1965d8d5a65e72c103b

Observation bb26dcbb-b4dd-4744-8451-210dee826867 · outbound

This paper cites Continuous control with deep reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Continuous control with deep reinforcement learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.295604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.295604Z digest=sha256:723ee176b9097db0bef46938ae17568e8f6a15f5680995ced347e5da9735d701

Observation acacfe4a-278d-49ad-9b0e-3b4e0ab72907 · outbound

This paper cites Aps: Active pretraining with successor features.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Aps: Active pretraining with successor features

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.241608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.301156Z digest=sha256:1e2e7a0803d9f32fd2db81858f9a5b963f01d3506ad21fb3f581d84c69fbaf6e

Observation 6e3ab296-d29d-45aa-8575-4f2845740779 · outbound

This paper cites Behavior from the void: Unsupervised active pre-training.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Behavior from the void: Unsupervised active pre-training

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.226391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.307110Z digest=sha256:ea4d2ceea098e12126697921a6637a3d4103cdaf7e9cd304e7be4673da5f1fbd

Observation 436e4d5e-93f0-458f-9091-e6ccdef8eb7a · outbound

This paper cites Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.313075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.313075Z digest=sha256:82767495d6e740b839907858eb6d84a222af27836c940fcd3179bb69f130b4fb

Observation 43d3f187-783d-4d0c-bf6d-39429be652df · outbound

This paper cites Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.210145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.318224Z digest=sha256:652de60aec8b94c5c36166f68779789256a6a42331420770816858b7b6e04c0d

Observation d6dd9ac4-0c13-4e8e-ba70-a58255ea0506 · outbound

This paper cites Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.323013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.323013Z digest=sha256:2b3505be4e3e40e48cd700cec1522ae5869ef8a175cf8abebd7bf2531a1960f6

Observation cab513be-a52c-4aa6-841e-378ec63bd8f9 · outbound

This paper cites Synthetic experience replay.Advances in Neural Information Processing Systems, 36, 2024.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Synthetic experience replay.Advances in Neural Information Processing Systems, 36, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.182147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.328698Z digest=sha256:8a3f457c4373c2f3dce64a7d82d606ab596271f2e01e3e450dd0c0e17279eb78

Observation e1fcfedc-9fa5-464d-813c-f463941d62e6 · outbound

This paper cites Efficient Online Reinforcement Learning for Diffusion Policy.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Efficient Online Reinforcement Learning for Diffusion Policy

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.333364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.333364Z digest=sha256:192f42510a272d0a435306a97083376fd5b5c8631d5c23aa5dd28182c73ac64d

Observation 4c5de78b-2d75-439a-83d5-9af270e71cc1 · outbound

This paper cites Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.339718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.339718Z digest=sha256:dc1d23f0e79f7d641e19b6d092b03952a4e707cde38fa8ae38dc04d94d851aaf

Observation ad88db33-c4b8-4834-933b-e230cbf6614d · outbound

This paper cites Curiosity-driven exploration via latent bayesian surprise.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Curiosity-driven exploration via latent bayesian surprise

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.166694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.344637Z digest=sha256:d08dbabad1aa399b897f25e2de0df367528e93aa693fa680a98598bfb5eaeb61

Observation e36c7829-ae27-43f0-8d40-9e40197bdf24 · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Lipschitz-constrained unsupervised skill discovery

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.151235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.349181Z digest=sha256:aea6715078d9ff0facf2f4b90f94691a497a8d32137de641233fdf0c7b1a29f2

Observation 25988992-dc49-4b2d-b2c8-54d202233e5e · outbound

This paper cites METRA: Scalable Unsupervised RL with Metric-Aware Abstraction.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning METRA: Scalable Unsupervised RL with Metric-Aware Abstraction

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.353849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.353849Z digest=sha256:1d2493e86196dc6306af4aba94c6a3a4a240627b037c981284afbf3351884dc4

Observation de555d34-5ec6-4044-b0ab-c4e20ecf47c2 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Curiosity-driven exploration by self-supervised prediction

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.135131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.358862Z digest=sha256:9e0b209bf92d4de934b25ce47740216e29c03139c61a5dc7425e753d81e71012

Observation 360a91dc-9430-4d88-ad63-b63cf128d0fd · outbound

This paper cites Self-supervised exploration via disagreement.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Self-supervised exploration via disagreement

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.118151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.363809Z digest=sha256:71b36da55f353b02f58dcff1c6817bcb48a2a9aeca4679b39afa629176b757e2

Observation b957b3af-87ce-4b1b-b3ef-1e63cc384755 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.368484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.368484Z digest=sha256:76f924387e93a5483faea9cb1a293862338a8214b4ad037bbb0fcdf8af80c72f

Observation 42042865-f0ce-403b-a6f5-260f47240610 · outbound

This paper cites Learning a Diffusion Model Policy from Rewards via Q-Score Matching.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Learning a Diffusion Model Policy from Rewards via Q-Score Matching

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.374992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.374992Z digest=sha256:982be10db22b3ad242a693ccb7eb36daff676dfe4dbccad619df1563089ee733

Observation 06931fde-8a40-45aa-87d0-7d1de4e1d46c · outbound

This paper cites Diffusion Policy Policy Optimization.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Policy Policy Optimization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.380367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.380367Z digest=sha256:3ae852f19c48846be3f0187336f71e8f14b3a0d30d75c4b8f4424a9adc307230

Observation 2707d75e-063d-4aec-be49-31b33b7b274d · outbound

This paper cites Photorealistic text-to- image diffusion models with deep language understanding.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Photorealistic text-to- image diffusion models with deep language understanding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.385361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.385361Z digest=sha256:40c5e3b42818fbaa6484fb5ed8c1fff9cc8eb5140d5de8e75bf12c3f457a61a1

Observation db3205d2-3abc-442f-97a3-48c031397268 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.390208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.390208Z digest=sha256:36d98aacedad13760a5120ed3956fede8db0c0c538a7468e63ccd98c15a86fd2

Observation cbbf8218-1a96-4e75-b11d-10589a6f930c · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Deep unsupervised learning using nonequilibrium thermodynamics

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.395721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.395721Z digest=sha256:43603b3286424356f59b43729d08f7287fe818befc199b2d1a24aeb3faca3327

Observation 9c1c4d90-8598-49f0-8f25-d4866102963a · outbound

This paper cites Denoising diffusion implicit models.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Denoising diffusion implicit models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.400244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.400244Z digest=sha256:a4409d872c5b140e78e7d5885e14aca459d9c183afaa6d61ef0554048360139e

Observation 413bb471-9086-4afb-904a-9a6070a93c95 · outbound

This paper cites Score-based generative modeling through stochastic differential equations.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Score-based generative modeling through stochastic differential equations

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.404894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.404894Z digest=sha256:058439b2bc10dc1224feb7dc1451d3708cbff71aed262d2ddf3a4e9d475faac7

Observation 0b3437f2-eb5b-4c53-89d8-2a520ae2f4b7 · outbound

This paper cites Reinforcement learning: An introduction.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Reinforcement learning: An introduction

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.409791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.409791Z digest=sha256:3cc8bc4e478cf79675615ecf15cc6f570ea0e7c00059fb66f345be4204916a35

Observation 55ee6be8-6836-40f9-a899-c5f22e3d4788 · outbound

This paper cites DeepMind Control Suite.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning DeepMind Control Suite

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.414534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.414534Z digest=sha256:b804418ce7dc237c5dae4f2c7cac27967ac6bc27cf05b619b9f376dc27f60574

Observation bce0a378-af0f-4c70-a587-9312f03ea012 · outbound

This paper cites Prioritized Generative Replay.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Prioritized Generative Replay

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.420119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.420119Z digest=sha256:f83eb63c833147f8a318a23f5f583d33aed1244faecfdcda4326d69e08b1f662

Observation 6594c9c2-0ac2-47fb-9fb0-2a70427a63b8 · outbound

This paper cites Diffusion policies as an expressive policy class for offline reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion policies as an expressive policy class for offline reinforcement learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.424969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.424969Z digest=sha256:d3f58a0acf489a9503553688426b1ad686b03d1454cbe3d7c54afcbb693eb801

Observation 0198e1c6-8411-45ae-951a-24aca56b6b25 · outbound

This paper cites A problem in geometric probability.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning A problem in geometric probability

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.039086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.429537Z digest=sha256:8ad83fdacfef8f0915769c6e4984b38072c1e14e64452e01c3a2c079f4aef185

Observation 5944f1a6-f8ea-44a9-90e7-e9a2249fdadb · outbound

This paper cites Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.434308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.434308Z digest=sha256:939092e904ba7103f8030e5192d38c258ad83a1cca3e6d15f515aa3e513a548e

Observation 3f30d08a-8ca5-447b-b349-8d05edb0bc5b · outbound

This paper cites Policy Representation via Diffusion Probability Model for Reinforcement Learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Policy Representation via Diffusion Probability Model for Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.438947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.438947Z digest=sha256:d19f82790e19c02927caa6faa7b9fe4284d61a88a19fa3b327233d44b7f4eaf8

Observation 78c1c545-d060-40a5-be76-0c7b9837b1b8 · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Behavior contrastive learning for unsupervised skill discovery

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.022950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.443810Z digest=sha256:c88784d03bb840fcce8d9e9e03eddd14dfa0dc2adfb2dccbe9486b978b2f570a

Observation a5ec26a0-6739-42d7-9d04-108188dd472f · outbound

This paper cites Peac: Unsupervised pre-training for cross-embodiment reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Peac: Unsupervised pre-training for cross-embodiment reinforcement learning

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:19.006929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.448752Z digest=sha256:78c39f78b2684f943a390cb15b22625111edb6c52f485d2ec92c36468df5078d

Observation 6ba8d291-d21f-4541-8de4-667a1bc26e33 · outbound

This paper cites Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.453588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.453588Z digest=sha256:01fa8bd6956dd846c1ec18ff0f870677a210195408179bc20072734bbf81be94

Observation 5376eb95-3c8f-4f54-84ab-3ed8f27c2477 · outbound

This paper cites Automatic intrinsic reward shaping for exploration in deep reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Automatic intrinsic reward shaping for exploration in deep reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:18.989867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.459069Z digest=sha256:cd73fbabb1d02c8cfac490253b1ca316beedd5469cb5698bd8e9fc3650428fe3

Observation 97a9345c-5642-4da6-94fc-5352cf84b381 · outbound

This paper cites EUCLID: Towards Efficient Unsupervised Reinforcement Learning with Multi-choice Dynamics Model.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning EUCLID: Towards Efficient Unsupervised Reinforcement Learning with Multi-choice Dynamics Model

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:18.464264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.464264Z digest=sha256:9f94328033f403cf2b2d8b8a4376a08baf42a5b6584bec3590deea1482c1d91a

Observation 7d7c98b1-be79-47c8-b37e-c988112b88fa · outbound

This paper cites A mixture of surprises for unsupervised reinforcement learning.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning A mixture of surprises for unsupervised reinforcement learning

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T13:21:18.973702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T13:21:18.469375Z digest=sha256:a3a335949a48c336a53a6f6d851e1f478cb63243840a19c1bb66c845069f49f6

Observation a33c2e8f-595f-448c-ba50-fecddc257fad · outbound

This paper cites Diffusion Models for Reinforcement Learning: A Survey.

Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Models for Reinforcement Learning: A Survey

Reference 70

Resolution
malformed identifier
no resolver link, observed 2026-08-08T13:21:18.474336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:18.474336Z digest=sha256:8ec1869131bfd41e73d1c70b5c4b47791545b070765e8919a637f51c0a9f8f23

Pith citing papers

No inbound Pith citation observations are available.