Pith. sign in

Paper Citation Record · LEDGER

Unsupervised Skill Discovery through Skill Regions Differentiation

As of 14 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2506.14420.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14420 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:27:00.722538Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 71f17265-6779-4c69-a3c3-c18f62a0d420 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,.

Unsupervised Skill Discovery through Skill Regions Differentiation A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.400282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.400282Z digest=sha256:3bf18d249b8ccec948040c511a58d58dfc690f685d799ccf21da35ed15d4d922

Observation ef1d3c3a-ae07-49e5-adc8-6e5fb3e25493 · outbound

This paper cites Mastering atari games with limited data,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering atari games with limited data,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.370724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.406553Z digest=sha256:6e214ea1ce39de27d2727ae02111c83a69f1c7c6949e4da6f0d13d48fc7d417a

Observation b2ec47ef-7373-48d0-a232-25be158a95b7 · outbound

This paper cites Continu- ous improvement of self-driving cars using dynamic confidence-aware reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Continu- ous improvement of self-driving cars using dynamic confidence-aware reinforcement learning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.360799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.410591Z digest=sha256:3bdda40520bafb428d3f61da6775d0fc89464fac6bb829a0f0116a0e9a1922a4

Observation 8394700b-fd26-46cf-b2d3-dd2159d76866 · outbound

This paper cites Uncertainty-aware model-based reinforce- ment learning: Methodology and application in autonomous driving,.

Unsupervised Skill Discovery through Skill Regions Differentiation Uncertainty-aware model-based reinforce- ment learning: Methodology and application in autonomous driving,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.350983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.415890Z digest=sha256:9a158b211cb5d3b83df15d3c2216782c4ac4656680533e4f48fc5366c7c0e5bd

Observation 704b35af-edb4-4fd9-9afc-82911f0acd5c · outbound

This paper cites Temporal difference learning for model predictive control,.

Unsupervised Skill Discovery through Skill Regions Differentiation Temporal difference learning for model predictive control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.340781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.421942Z digest=sha256:56cd421c2ce26bf1be566cb71c9a8ba246fbf6a56287c35072904308bd2849d7

Observation ed24a445-4e3a-4971-a598-29b233b89200 · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning robust perceptive locomotion for quadrupedal robots in the wild,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.330608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.426007Z digest=sha256:952eef4ef086fc8e88512d17d37673a9201a04c1748e0a2e0295dee0ed1e502d

Observation 8dde356a-dbf6-43df-8776-353702a37f9c · outbound

This paper cites Relay hindsight experience replay: Self-guided continual reinforcement learning for sequential object manipulation tasks with sparse rewards,.

Unsupervised Skill Discovery through Skill Regions Differentiation Relay hindsight experience replay: Self-guided continual reinforcement learning for sequential object manipulation tasks with sparse rewards,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.430953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.430953Z digest=sha256:96847c875a702ae62c2e5cc8ba9309d159df4593178c0809e37053aa3871e0fd

Observation e58e16ff-5cdf-4271-b646-8e358a74f9e9 · outbound

This paper cites Reward design with language models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reward design with language models,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.314030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.434350Z digest=sha256:c15fb248c99f0d7e097e1a7942a813e53825abb17a22cdf5fec2ea56dd2c85ea

Observation d4c1edc9-f530-46c2-bf7b-c3e216841b4f · outbound

This paper cites Maniskill2: A unified benchmark for generalizable manipulation skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Maniskill2: A unified benchmark for generalizable manipulation skills,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.303638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.438282Z digest=sha256:1f638cf8eb98c7589f8105317fb5c6831202409f821cb13e0434b4d624bd5bd3

Observation 619f9568-8743-48ab-b70a-aa94a8152069 · outbound

This paper cites Pre-trained models: Past, present and future,.

Unsupervised Skill Discovery through Skill Regions Differentiation Pre-trained models: Past, present and future,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.293541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.443129Z digest=sha256:05859da8bdefda51068ab7a23396e0cef58fc4966614ccfbd94a645976b2985f

Observation 4fd48582-3036-4fbe-860a-e6ca0ba79d71 · outbound

This paper cites GPT-4 Technical Report.

Unsupervised Skill Discovery through Skill Regions Differentiation GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.448023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.448023Z digest=sha256:2c3fd45bce991910709630771ac269d779558d9167556c3be4d2043d5f58ac8e

Observation a57c93fb-c2b9-415d-9369-dde7bd856de2 · outbound

This paper cites Training language models to follow instructions with human feedback,.

Unsupervised Skill Discovery through Skill Regions Differentiation Training language models to follow instructions with human feedback,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.452438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.452438Z digest=sha256:b93d9ed363a924868448e3ac4933aca7e131544ab3c7f6d3c5c17f595ee98965

Observation 6011ef50-a54c-4a95-9f8a-97e8b97c3e0f · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Unsupervised Skill Discovery through Skill Regions Differentiation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.456970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.456970Z digest=sha256:95a61f2ec1c54a9619902016aa615ad99c64a2fd776a8cf8e9fc9fe05f7f2f4d

Observation 226fe321-0ecf-4b4d-aff0-bbe46e7b5c6e · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Unsupervised Skill Discovery through Skill Regions Differentiation Masked au- toencoders are scalable vision learners,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.461759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.461759Z digest=sha256:73b3cb434f6e34b5156d0c6ede0ea36aa6371caeb02fc41e739ef0916e2a45c6

Observation fe02bf74-1ab5-4135-b72a-c480516daa08 · outbound

This paper cites V-JEPA: Latent video prediction for visual representation learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation V-JEPA: Latent video prediction for visual representation learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.268456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.465949Z digest=sha256:1fc5fe327cd7b1d0b75bc551152573b1fcf865f2f4441f6a5437c72c047505fb

Observation b20b188b-cf5e-4a2c-8a52-136036aad63f · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Bootstrap your own latent-a new approach to self-supervised learning,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.472353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.472353Z digest=sha256:cdd99d3d059be6247c51174b0cd16b3d98a92e376f3cc6dc1cebc56db8401d40

Observation 30b3293a-b4c1-4a8c-83b1-2ed348eb81f4 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

Unsupervised Skill Discovery through Skill Regions Differentiation Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.250413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.478396Z digest=sha256:390f9e41d5cf8aeec3300c37301c8ff191d18035ec93db9458a0ffd786e21d60

Observation 220c38ac-5d7d-4b90-bcd2-44cddea60b37 · outbound

This paper cites R3m: A universal visual representation for robot manipulation,.

Unsupervised Skill Discovery through Skill Regions Differentiation R3m: A universal visual representation for robot manipulation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.240433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.485375Z digest=sha256:e9f960c737396b4e9b4961b2f6171c39df2e879cff45596a8122c988d0b15122

Observation 36d4a9dc-19eb-4d33-801a-cf0dfc0a383f · outbound

This paper cites URLB: Unsupervised reinforcement learning benchmark,.

Unsupervised Skill Discovery through Skill Regions Differentiation URLB: Unsupervised reinforcement learning benchmark,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.230131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.489329Z digest=sha256:599d61512caba8cefc0e90678128a5494eafdd50b800b6e95d384a87645fa9e6

Observation 904361d9-e69c-40ef-bd35-575853178d1d · outbound

This paper cites Variational Intrinsic Control.

Unsupervised Skill Discovery through Skill Regions Differentiation Variational Intrinsic Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.497698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.497698Z digest=sha256:3e59c39116c169ad879c9e4cbc3b57c112db7b2b8a19476b99b11eda509cf569

Observation 7b6e1eed-5803-4e2e-8b8c-5842aef46ac9 · outbound

This paper cites Behavior from the void: Unsupervised active pre-training,.

Unsupervised Skill Discovery through Skill Regions Differentiation Behavior from the void: Unsupervised active pre-training,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.219172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.501331Z digest=sha256:93c3b1057f56d5744e3c51f8e4b6567b5f42fd116eba4ff4d80af7e6d657d21f

Observation a619baad-5a87-4765-9ed0-96b73d1ae417 · outbound

This paper cites Understanding the limitations of variational mutual information estimators,.

Unsupervised Skill Discovery through Skill Regions Differentiation Understanding the limitations of variational mutual information estimators,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.207390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.505838Z digest=sha256:c2eec8ed2116af0aba865dcb7d7f4d0bf19b6c181d5afd07d43e20d5e64a6875

Observation 310b62bc-be25-4cdf-96d8-474153cd3b5c · outbound

This paper cites Diversity is all you need: Learning skills without a reward function,.

Unsupervised Skill Discovery through Skill Regions Differentiation Diversity is all you need: Learning skills without a reward function,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.196242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.510281Z digest=sha256:ba952e9eadd918a1594f8825a4acda310666530577024f8e1308a25bfe43c200

Observation 0eaa15a3-04d2-4e2a-b9e0-2a1cf9db9cdc · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Behavior contrastive learning for unsupervised skill discovery,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.185751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.514388Z digest=sha256:5df035b392d65e187d6c2ed507df2559047e826a22d77246901346a943214ec1

Observation 5f7cea8d-1147-4b16-82f1-665d802b6b56 · outbound

This paper cites METRA: Scalable unsupervised RL with metric-aware abstraction,.

Unsupervised Skill Discovery through Skill Regions Differentiation METRA: Scalable unsupervised RL with metric-aware abstraction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.174666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.520823Z digest=sha256:f53f0f5ecffd0439932ca0db468ba6ec0baf0a63918a8005ce7a129d3a85d38e

Observation 455f1e32-4fc6-44e9-a00c-3ae7ae9e7605 · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Lipschitz-constrained unsupervised skill discovery,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.530003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.530003Z digest=sha256:b1a5efc12c4dab76f822db2dca7d721d163c321e8a7730520e5481004f06d6df

Observation 8e11a816-fa34-4e9e-9f6b-eb50d1a3912a · outbound

This paper cites Controllability-aware unsuper- vised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Controllability-aware unsuper- vised skill discovery,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.157930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.540883Z digest=sha256:5d9167263fd28345b4bd0c2d04b9e6931e9007badb76d672b2247c673793c06b

Observation 0151c8b4-a341-4e52-add7-ba2c19c31fe1 · outbound

This paper cites Unsupervised reinforcement learning with contrastive intrinsic control,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised reinforcement learning with contrastive intrinsic control,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.146584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.554184Z digest=sha256:ad0bb564a5b5bab274734ff366bd7e14194634e222b9a73bcf56bec6076ba53e

Observation 9834ffa6-f4e4-4a50-bde2-7cc8dff66686 · outbound

This paper cites Mastering the unsupervised reinforce- ment learning benchmark from pixels,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering the unsupervised reinforce- ment learning benchmark from pixels,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.135541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.565679Z digest=sha256:ba4efa00ea60e45a416de978b704aea6f3de33440f94a5c8b246496634ca63f7

Observation eb901812-b8f3-4efb-8a55-47a2b42fc20c · outbound

This paper cites Auto-Encoding Variational Bayes.

Unsupervised Skill Discovery through Skill Regions Differentiation Auto-Encoding Variational Bayes

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.580787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.580787Z digest=sha256:5448551ed7cbd3b2dbd5d1f3b8222a1867e331c675cc4eb2678e0335ff154240

Observation 657529dc-3a0d-4c4c-b252-2c7bf0b89cc7 · outbound

This paper cites An introduction to variational autoencoders,.

Unsupervised Skill Discovery through Skill Regions Differentiation An introduction to variational autoencoders,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.594940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.594940Z digest=sha256:5f36203c25f82176c08825ac53d9c7640ba093278776d3661fde6949671066fc

Observation bf395112-8f86-4dee-8e21-659391494978 · outbound

This paper cites Multi-task reinforcement learning with soft modularization,.

Unsupervised Skill Discovery through Skill Regions Differentiation Multi-task reinforcement learning with soft modularization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.118194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.598765Z digest=sha256:31372c6a5687bcfaab5ec8f718e275732f7109121eeba90e68be8f404f2ca990

Observation 83799067-9158-4cf0-8018-88614f6ce87f · outbound

This paper cites Near-bayesian exploration in polynomial time,.

Unsupervised Skill Discovery through Skill Regions Differentiation Near-bayesian exploration in polynomial time,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.107853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.606260Z digest=sha256:73c41bc7c3a2451689ec6600c2f8e96a2adc794fade51022055e40c7ccd5cce8

Observation 772d9560-1c81-45ba-bc53-39168129a0c9 · outbound

This paper cites An analysis of model-based interval estimation for markov decision processes,.

Unsupervised Skill Discovery through Skill Regions Differentiation An analysis of model-based interval estimation for markov decision processes,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.096620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.609774Z digest=sha256:1242ecdbf119d702f7e8a2fa037a6e71a79cfbc67e7046d50b1a8f2457d42f72

Observation 2809328a-2103-4dae-befb-55cdcac852fd · outbound

This paper cites Unifying count-based exploration and intrinsic motivation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unifying count-based exploration and intrinsic motivation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.087603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.612633Z digest=sha256:b4d83eef3b89913051e056d1e81a2d33805b0ae16bae0bef3d69eda1ce8199a2

Observation 072ed9ba-5c18-4643-97a3-a90c2c7dcb13 · outbound

This paper cites Count-based exploration with neural density models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Count-based exploration with neural density models,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.078190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.616640Z digest=sha256:2e921d504b91f497b88c0331f78bab084b6f16eecac73b48b78932d8600e5d83

Observation 530eb6f2-b55d-4f03-9206-a4563300a6aa · outbound

This paper cites Dynamics- aware unsupervised discovery of skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Dynamics- aware unsupervised discovery of skills,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.068018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.619907Z digest=sha256:f513e391f3bd0f6f3f4418f47eb5f76390b71a6bbb423b95dd541b2adc876d57

Observation cfa7a766-a5eb-47db-a842-e327c7ea0c66 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Explore, discover and learn: Unsupervised discovery of state-covering skills,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.058443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.623315Z digest=sha256:e38efbcffe9d7ff5eb3a5181c50570088b494ea171d7b6573fd7339715f91484

Observation 4a3bb0fb-7ccb-49ae-b38b-d94623953078 · outbound

This paper cites Unsupervised skill discovery via recurrent skill training,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised skill discovery via recurrent skill training,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.048859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.626404Z digest=sha256:8d5ff3cfc126218ddddb8e3360c0182f987a961248c5a70dec81a92e4396b214

Observation 03fd7f24-e87f-48cf-9479-f453ef079ad9 · outbound

This paper cites Learning to discover skills with guidance,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning to discover skills with guidance,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.038583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.629803Z digest=sha256:9649386123b4bef309af7aa8172cd07c60d05f3e5eb3336c7260e00782570a41

Observation 5bc42a72-5855-459d-98d9-059d07f1d69e · outbound

This paper cites Aps: Active pretraining with successor features,.

Unsupervised Skill Discovery through Skill Regions Differentiation Aps: Active pretraining with successor features,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.028714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.633081Z digest=sha256:2bdee2c8919a096abbc37dd94ee17f4f65ee26bc3018f476c338e603832b30ef

Observation da26a86e-3728-495a-9390-35aa84fa09de · outbound

This paper cites Efficient exploration via state marginal matching,.

Unsupervised Skill Discovery through Skill Regions Differentiation Efficient exploration via state marginal matching,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.017875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.636865Z digest=sha256:f112335066bc3650da04fbb45807d520f70a8bfb2aa799de5f0c046beb4ee6f9

Observation f32b9b01-64de-403f-95ca-bded11c4e539 · outbound

This paper cites Learning more skills through optimistic exploration,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning more skills through optimistic exploration,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.997823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.644518Z digest=sha256:4fba4cd2c97c3b3f977861660024f9fdbfc7f62e9aa2cbcb0faed13d8675d708

Observation f6728e22-9bc1-4fde-bf6b-b38aeb7858cf · outbound

This paper cites Choreographer: Learning and adapting skills in imagination,.

Unsupervised Skill Discovery through Skill Regions Differentiation Choreographer: Learning and adapting skills in imagination,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.987631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.647901Z digest=sha256:8d24ab5df7a9c512a6bdf95174db2b50698830e37766de868e5e4cea3b52893d

Observation 57370c5e-b6cf-4ac9-a791-2a0f30a6ab51 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction,.

Unsupervised Skill Discovery through Skill Regions Differentiation Curiosity-driven exploration by self-supervised prediction,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.978763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.651425Z digest=sha256:dd9cf4e1db5a9cf35a64b85ee4551b1e344fba80d8d78f74a3a32f50a2d61554

Observation cd209605-fe50-4286-bb76-0bd9626397f5 · outbound

This paper cites Self-supervised exploration via disagreement,.

Unsupervised Skill Discovery through Skill Regions Differentiation Self-supervised exploration via disagreement,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.968809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.655593Z digest=sha256:bdd67a50e92385f1bd8cf067034023bedb0548e1ee26430d1715674f436dffd2

Observation 6f595c2b-5c8a-4abe-95f3-e0ea93083b26 · outbound

This paper cites Exploration by random network distillation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Exploration by random network distillation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.959226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.659560Z digest=sha256:65cadb022fbefc5f987a14d5f6a8b89ce75c4e369515c6e56788fa6efb3e6b66

Observation e160b572-259a-4e80-b456-0ce9ac71b60d · outbound

This paper cites Reinforcement learn- ing with prototypical representations,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reinforcement learn- ing with prototypical representations,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.949840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.662915Z digest=sha256:adffef0f4e8ebc0f6f84b930863957310884efee4ee7fe8cafd883cda0971b35

Observation 60d6b8f5-1755-43fa-be8b-c905890209f9 · outbound

This paper cites Unsupervised Skill-Discovery and Skill-Learning in Minecraft.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised Skill-Discovery and Skill-Learning in Minecraft

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.666160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.666160Z digest=sha256:fa1a66c6dda6c47b758f8079b4d2601e01e43af6d5a36636c2d615f51ab17b8d

Observation 0174ae71-1d5d-4b66-b325-c726bebd3b3a · outbound

This paper cites Rethinking mutual information for language conditioned skill discovery on imitation learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Rethinking mutual information for language conditioned skill discovery on imitation learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.940307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.670115Z digest=sha256:d6b3b629b67b0cf4b316e21d3033aebfcf0c5360e0c78b1acb15d001f1cd2cab

Observation 87d5dae8-0c73-47fd-adea-1a04ce1e5fd5 · outbound

This paper cites Robust policy learning via offline skill diffusion,.

Unsupervised Skill Discovery through Skill Regions Differentiation Robust policy learning via offline skill diffusion,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.930800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.673501Z digest=sha256:189b2d0823ec6faff016ff7ee5b2677de1784d2bff95029d0f85b148301232fb

Observation 532129b3-6699-4fdc-a237-47c38bbbd6c6 · outbound

This paper cites EUCLID: Towards efficient unsupervised reinforcement learning with multi-choice dynamics model,.

Unsupervised Skill Discovery through Skill Regions Differentiation EUCLID: Towards efficient unsupervised reinforcement learning with multi-choice dynamics model,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.920765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.676484Z digest=sha256:ea80f188861aa2b8ba0f38d06f311aa8703a6dd8af8a83352be54126e0277138

Observation 9b7851eb-dfa8-4001-93cb-37191848cf60 · outbound

This paper cites Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning.

Unsupervised Skill Discovery through Skill Regions Differentiation Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.680044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.680044Z digest=sha256:a9ff15f07c79475a3b611a94b4badb7cae9a5e977350a3379df279c6578e4567

Observation c0d40dd6-4f94-410f-9d1e-31652ed883fa · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice,.

Unsupervised Skill Discovery through Skill Regions Differentiation Deep reinforcement learning at the edge of the statistical precipice,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.910096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.683746Z digest=sha256:c2e0b8e4494f751c498030e55b0482c79e8ad508a16ec4e99ff32549c8f7fe0b

Observation 66b480bf-5b9a-445a-93bd-0ddbd0d3441a · outbound

This paper cites DeepMind Control Suite.

Unsupervised Skill Discovery through Skill Regions Differentiation DeepMind Control Suite

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.688031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.688031Z digest=sha256:87921eccc562922e28274e93358766757e33e8c57d0b297c4876f2f36c54b0da

Observation a9f30ed3-f669-4d5f-b7d8-e7b1c2a40bed · outbound

This paper cites Continuous control with deep reinforcement learning.

Unsupervised Skill Discovery through Skill Regions Differentiation Continuous control with deep reinforcement learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.691720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.691720Z digest=sha256:87b45008456ca1f927dbd10cc2b255de92fb3d1fc38949f88c467640f64deff0

Observation 938e136d-270b-44d4-879d-efd40bcb47a9 · outbound

This paper cites Mastering visual continuous control: Improved data-augmented reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering visual continuous control: Improved data-augmented reinforcement learning,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.900348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.695397Z digest=sha256:96581e2710486bf691cfd86f0f31b4931c6a677d1c8f110d69c99645ec68a015

Observation c14da304-2640-44aa-a0a0-67dbafd95f84 · outbound

This paper cites Mastering atari with discrete world models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering atari with discrete world models,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.890594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.699020Z digest=sha256:43527e2f8a39653ed398b843dbdf14e8b33bfaf8778fc34b7b25ce54125b9c02

Observation 05d9e23b-2958-4fc4-b1d1-809d2a638e31 · outbound

This paper cites Provably efficient rein- forcement learning with linear function approximation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Provably efficient rein- forcement learning with linear function approximation,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.880897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.702736Z digest=sha256:38428784892900292f20cd18748b45024376039c0e36cb3bcfe7ddf1680b0cd2

Observation cf051aba-497b-4290-b167-116d05c26cdd · outbound

This paper cites Reward-free model-based reinforcement learning with linear function approximation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reward-free model-based reinforcement learning with linear function approximation,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.871448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.707318Z digest=sha256:2b448f673ec7f99316e89062e9ddef74d69577902b940959f6141d4e0975d1c4

Observation ae0bbdbe-c31f-4a3d-8441-e11393c504ed · outbound

This paper cites Deep variational information bottleneck,.

Unsupervised Skill Discovery through Skill Regions Differentiation Deep variational information bottleneck,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.711258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.711258Z digest=sha256:807a17909caf15584fd8d95715851e0a02e7bd402025189ca3ad1e59d5cec4a5

Observation 66625992-49c9-4b03-abe3-5d4fee27eec8 · outbound

This paper cites Dynamic bottleneck for robust self-supervised exploration,.

Unsupervised Skill Discovery through Skill Regions Differentiation Dynamic bottleneck for robust self-supervised exploration,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.854381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.715141Z digest=sha256:10ab1f7e2065533dff849456f050f730b3a97468f460bcbb6be172cae38c9e22

Observation 17e3e0cf-fcfd-4ab5-a163-7e6f6baa304e · outbound

This paper cites Provably efficient exploration in policy optimization,.

Unsupervised Skill Discovery through Skill Regions Differentiation Provably efficient exploration in policy optimization,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.844208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.718687Z digest=sha256:997536036534a5224c2a0cd0a8a9f1fcd6d18fe4c6e7d0e02dda9364efd45462

Observation d95da145-63a3-467a-88a2-5c8a91764329 · outbound

This paper cites Logarithmic online regret bounds for undis- counted reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Logarithmic online regret bounds for undis- counted reinforcement learning,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.832856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.722538Z digest=sha256:c80ce78026cc107362690f8a9a415b33ab237fec544443eb3088092dc067dddf

Observation 591aa826-3252-4443-a02e-6acff41e581c · outbound

This paper cites Available: https://openreview.net/forum?id=Hkla1eHFvS.

Unsupervised Skill Discovery through Skill Regions Differentiation Available: https://openreview.net/forum?id=Hkla1eHFvS

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.008510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T00:27:00.640914Z digest=sha256:7f8747e7edc2bb0ae6a5f739f27d4e4a210507f109c74efa3a51455fc62baa52

Pith citing papers

No inbound Pith citation observations are available.