Pith. sign in

Paper Citation Record · LEDGER

Unsupervised Skill Discovery through Skill Regions Differentiation

As of 19 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2506.14420.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14420 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:27:00.722538Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 71f17265-6779-4c69-a3c3-c18f62a0d420 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,.

Unsupervised Skill Discovery through Skill Regions Differentiation A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.400282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.400282Z digest=sha256:8f324b7ca61332064ba391a1ed7d9354cf5083ee618f3d82df1fa55924469af7

Observation ef1d3c3a-ae07-49e5-adc8-6e5fb3e25493 · outbound

This paper cites Mastering atari games with limited data,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering atari games with limited data,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.370724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.406553Z digest=sha256:83e9dc064c3b71499e9b35f7ddbbac5da0c5650522c4e7d2796c029b2d786903

Observation b2ec47ef-7373-48d0-a232-25be158a95b7 · outbound

This paper cites Continu- ous improvement of self-driving cars using dynamic confidence-aware reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Continu- ous improvement of self-driving cars using dynamic confidence-aware reinforcement learning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.360799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.410591Z digest=sha256:e6532d21ea04f8c2757ef978ee71fe368b1d20ac754b1b059b04981ce6c26e00

Observation 8394700b-fd26-46cf-b2d3-dd2159d76866 · outbound

This paper cites Uncertainty-aware model-based reinforce- ment learning: Methodology and application in autonomous driving,.

Unsupervised Skill Discovery through Skill Regions Differentiation Uncertainty-aware model-based reinforce- ment learning: Methodology and application in autonomous driving,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.350983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.415890Z digest=sha256:313465303b9d8e3650856dbd2ecaedf49d53321fedf5ef59f95320d3bc885e45

Observation 704b35af-edb4-4fd9-9afc-82911f0acd5c · outbound

This paper cites Temporal difference learning for model predictive control,.

Unsupervised Skill Discovery through Skill Regions Differentiation Temporal difference learning for model predictive control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.340781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.421942Z digest=sha256:4bacc442ae56b1e03ad938f5cb4bf031e6a65e862a01be5aa7ac4221187776d7

Observation ed24a445-4e3a-4971-a598-29b233b89200 · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning robust perceptive locomotion for quadrupedal robots in the wild,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.330608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.426007Z digest=sha256:57200c453d69ffa07af050842375702ec64a7280d2cd537bf0ed277ad12416f9

Observation 8dde356a-dbf6-43df-8776-353702a37f9c · outbound

This paper cites Relay hindsight experience replay: Self-guided continual reinforcement learning for sequential object manipulation tasks with sparse rewards,.

Unsupervised Skill Discovery through Skill Regions Differentiation Relay hindsight experience replay: Self-guided continual reinforcement learning for sequential object manipulation tasks with sparse rewards,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.430953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.430953Z digest=sha256:0eece0614276802d3b9e22b4a8373548f8aae8e4e50d609aa8095df5ed10e658

Observation e58e16ff-5cdf-4271-b646-8e358a74f9e9 · outbound

This paper cites Reward design with language models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reward design with language models,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.314030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.434350Z digest=sha256:0de90170ae29961321adbccc3094e2fbbd4de2982be1abd64ea3e851ffe6056a

Observation d4c1edc9-f530-46c2-bf7b-c3e216841b4f · outbound

This paper cites Maniskill2: A unified benchmark for generalizable manipulation skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Maniskill2: A unified benchmark for generalizable manipulation skills,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.303638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.438282Z digest=sha256:e4bf21f4045fb300c0e26d30914ac073255614986a462087b528cbcafd115439

Observation 619f9568-8743-48ab-b70a-aa94a8152069 · outbound

This paper cites Pre-trained models: Past, present and future,.

Unsupervised Skill Discovery through Skill Regions Differentiation Pre-trained models: Past, present and future,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.293541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.443129Z digest=sha256:8da6ee4ca9390b5426eac3d0c6f62c9e7d12b2af7e579c168c2eb6ba76bdcf2d

Observation 4fd48582-3036-4fbe-860a-e6ca0ba79d71 · outbound

This paper cites GPT-4 Technical Report.

Unsupervised Skill Discovery through Skill Regions Differentiation GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.448023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.448023Z digest=sha256:6ef94fcad0dd2bda0454664211d0ddd2462a0a968918cc521219d1104b3b747c

Observation a57c93fb-c2b9-415d-9369-dde7bd856de2 · outbound

This paper cites Training language models to follow instructions with human feedback,.

Unsupervised Skill Discovery through Skill Regions Differentiation Training language models to follow instructions with human feedback,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.452438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.452438Z digest=sha256:e6a437ad1508bdfaa6fd7e722819c1ff2ba7bb05d6bf5880ab9417ac9e368ef8

Observation 6011ef50-a54c-4a95-9f8a-97e8b97c3e0f · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Unsupervised Skill Discovery through Skill Regions Differentiation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.456970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.456970Z digest=sha256:1470745b0e820f45ee98fbbfd2ac563a9b18f780c7b5d600de86de4784ec1899

Observation 226fe321-0ecf-4b4d-aff0-bbe46e7b5c6e · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Unsupervised Skill Discovery through Skill Regions Differentiation Masked au- toencoders are scalable vision learners,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.461759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.461759Z digest=sha256:c351f47129cd8d869170d08f310c74f2668e5d3e84337f48fbe6660d0f69bf40

Observation fe02bf74-1ab5-4135-b72a-c480516daa08 · outbound

This paper cites V-JEPA: Latent video prediction for visual representation learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation V-JEPA: Latent video prediction for visual representation learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.268456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.465949Z digest=sha256:aad54135f5546836aaf592ed3662d2c7db46b8ac1700c0a809a31f02eb43761e

Observation b20b188b-cf5e-4a2c-8a52-136036aad63f · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Bootstrap your own latent-a new approach to self-supervised learning,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.472353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.472353Z digest=sha256:0ba8770fe41543116928deb0571b15b76b834c745fe2cb5e5735015885eda4c9

Observation 30b3293a-b4c1-4a8c-83b1-2ed348eb81f4 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

Unsupervised Skill Discovery through Skill Regions Differentiation Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.250413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.478396Z digest=sha256:f805f3c221f21312aad65af0ebb77f81534a61b3566bf755f0987a0ebc74653c

Observation 220c38ac-5d7d-4b90-bcd2-44cddea60b37 · outbound

This paper cites R3m: A universal visual representation for robot manipulation,.

Unsupervised Skill Discovery through Skill Regions Differentiation R3m: A universal visual representation for robot manipulation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.240433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.485375Z digest=sha256:39b4e686f2c3da5dd80266a3ca35e04adae98e4ee0f588bd1559924ae7c5b7f1

Observation 36d4a9dc-19eb-4d33-801a-cf0dfc0a383f · outbound

This paper cites URLB: Unsupervised reinforcement learning benchmark,.

Unsupervised Skill Discovery through Skill Regions Differentiation URLB: Unsupervised reinforcement learning benchmark,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.230131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.489329Z digest=sha256:1cc88c75b9c82dcb699c2c902f3494da3470e5132d57113e432dfc496778f350

Observation 904361d9-e69c-40ef-bd35-575853178d1d · outbound

This paper cites Variational Intrinsic Control.

Unsupervised Skill Discovery through Skill Regions Differentiation Variational Intrinsic Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.497698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.497698Z digest=sha256:bd019d70dd5774855825f9b68839446762272aad8fd1f4ba1d690f5970c36ad2

Observation 7b6e1eed-5803-4e2e-8b8c-5842aef46ac9 · outbound

This paper cites Behavior from the void: Unsupervised active pre-training,.

Unsupervised Skill Discovery through Skill Regions Differentiation Behavior from the void: Unsupervised active pre-training,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.219172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.501331Z digest=sha256:cce21300d496c199c1876b87e80eee703a31331fb9c81a56adf122645919d7db

Observation a619baad-5a87-4765-9ed0-96b73d1ae417 · outbound

This paper cites Understanding the limitations of variational mutual information estimators,.

Unsupervised Skill Discovery through Skill Regions Differentiation Understanding the limitations of variational mutual information estimators,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.207390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.505838Z digest=sha256:acba263503695ea3b17c795c7b26a7d88bf2e9b32394ce31d3128ad248d42307

Observation 310b62bc-be25-4cdf-96d8-474153cd3b5c · outbound

This paper cites Diversity is all you need: Learning skills without a reward function,.

Unsupervised Skill Discovery through Skill Regions Differentiation Diversity is all you need: Learning skills without a reward function,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.196242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.510281Z digest=sha256:1ddd4dbb2a3c2118c1f030a64e759bd1aac3dbb30c72df3335d9358003ac1a93

Observation 0eaa15a3-04d2-4e2a-b9e0-2a1cf9db9cdc · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Behavior contrastive learning for unsupervised skill discovery,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.185751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.514388Z digest=sha256:30b83fccf2c0f552c299ee9b6f8531714c3052c04b03ac350978b27af4d38c6c

Observation 5f7cea8d-1147-4b16-82f1-665d802b6b56 · outbound

This paper cites METRA: Scalable unsupervised RL with metric-aware abstraction,.

Unsupervised Skill Discovery through Skill Regions Differentiation METRA: Scalable unsupervised RL with metric-aware abstraction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.174666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.520823Z digest=sha256:62875527037619f57c8725ec309d436afdc51b9f5403c1d720368d4eea6c0136

Observation 455f1e32-4fc6-44e9-a00c-3ae7ae9e7605 · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Lipschitz-constrained unsupervised skill discovery,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.530003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.530003Z digest=sha256:2da068deb460689d2d40d80272f171d5c3bc783634ee3d7af7fbaa41ecab7805

Observation 8e11a816-fa34-4e9e-9f6b-eb50d1a3912a · outbound

This paper cites Controllability-aware unsuper- vised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Controllability-aware unsuper- vised skill discovery,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.157930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.540883Z digest=sha256:c9816a59ff389069b56e1f8888a6265ab5c10d10e570dbd7ca7a225c6a2ccd48

Observation 0151c8b4-a341-4e52-add7-ba2c19c31fe1 · outbound

This paper cites Unsupervised reinforcement learning with contrastive intrinsic control,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised reinforcement learning with contrastive intrinsic control,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.146584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.554184Z digest=sha256:ba777159a01da98a34fa5cde4d32d781b5d7cd6422a8fc4c95df42762bcb955f

Observation 9834ffa6-f4e4-4a50-bde2-7cc8dff66686 · outbound

This paper cites Mastering the unsupervised reinforce- ment learning benchmark from pixels,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering the unsupervised reinforce- ment learning benchmark from pixels,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.135541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.565679Z digest=sha256:81f3248d086a3fc2df1487a9e897a2b3857356236d225888259999bf1c0f16af

Observation eb901812-b8f3-4efb-8a55-47a2b42fc20c · outbound

This paper cites Auto-Encoding Variational Bayes.

Unsupervised Skill Discovery through Skill Regions Differentiation Auto-Encoding Variational Bayes

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.580787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.580787Z digest=sha256:8ae113e21fbfea6d6dfc5bc9b30d708a348f8f71a6ba7dc8ea0a19c938c4179a

Observation 657529dc-3a0d-4c4c-b252-2c7bf0b89cc7 · outbound

This paper cites An introduction to variational autoencoders,.

Unsupervised Skill Discovery through Skill Regions Differentiation An introduction to variational autoencoders,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.594940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.594940Z digest=sha256:5823f268e3ab31659930ef2a7089eb18beb250e808a6df5f29ec15ec69a954d2

Observation bf395112-8f86-4dee-8e21-659391494978 · outbound

This paper cites Multi-task reinforcement learning with soft modularization,.

Unsupervised Skill Discovery through Skill Regions Differentiation Multi-task reinforcement learning with soft modularization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.118194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.598765Z digest=sha256:13bcffb1258309c8d92cb8f97dc403b09cb56b00a958d13fb53c94cd771d58e2

Observation 83799067-9158-4cf0-8018-88614f6ce87f · outbound

This paper cites Near-bayesian exploration in polynomial time,.

Unsupervised Skill Discovery through Skill Regions Differentiation Near-bayesian exploration in polynomial time,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.107853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.606260Z digest=sha256:b23d6f82e3f57e38042c91bdd833c867e6d582f09042d89608f6cead18d47f2a

Observation 772d9560-1c81-45ba-bc53-39168129a0c9 · outbound

This paper cites An analysis of model-based interval estimation for markov decision processes,.

Unsupervised Skill Discovery through Skill Regions Differentiation An analysis of model-based interval estimation for markov decision processes,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.096620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.609774Z digest=sha256:b735931b2f5d141e7479b3869cec747fa2a3bfa4d86ecff1a46c17848693b74a

Observation 2809328a-2103-4dae-befb-55cdcac852fd · outbound

This paper cites Unifying count-based exploration and intrinsic motivation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unifying count-based exploration and intrinsic motivation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.087603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.612633Z digest=sha256:87a7aed4c6722a1f9837e466e8c18f913970e556ca2bc23dc5e3a9751f2e5e2f

Observation 072ed9ba-5c18-4643-97a3-a90c2c7dcb13 · outbound

This paper cites Count-based exploration with neural density models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Count-based exploration with neural density models,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.078190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.616640Z digest=sha256:fc16fc322f036eaf3486725b3bf6bebf4ce8756c1fcd399aefc3f83015698a7d

Observation 530eb6f2-b55d-4f03-9206-a4563300a6aa · outbound

This paper cites Dynamics- aware unsupervised discovery of skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Dynamics- aware unsupervised discovery of skills,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.068018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.619907Z digest=sha256:b1227a01cba8196c6584e2c0f328e7013d85ea60185dacd3d298be4c0e7fee21

Observation cfa7a766-a5eb-47db-a842-e327c7ea0c66 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Explore, discover and learn: Unsupervised discovery of state-covering skills,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.058443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.623315Z digest=sha256:c819c80a66c73f62859bb43f71d36514158f4209fc94236e5b2c24a331d044f2

Observation 4a3bb0fb-7ccb-49ae-b38b-d94623953078 · outbound

This paper cites Unsupervised skill discovery via recurrent skill training,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised skill discovery via recurrent skill training,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.048859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.626404Z digest=sha256:aef32623ff2c2ba01d8e040444b2aa426441c0facba6c649afc708902c5261c4

Observation 03fd7f24-e87f-48cf-9479-f453ef079ad9 · outbound

This paper cites Learning to discover skills with guidance,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning to discover skills with guidance,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.038583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.629803Z digest=sha256:140e188ae3f1e1f7d422ec1641219ac2b2835222fd560b9fc7b553235786f9d8

Observation 5bc42a72-5855-459d-98d9-059d07f1d69e · outbound

This paper cites Aps: Active pretraining with successor features,.

Unsupervised Skill Discovery through Skill Regions Differentiation Aps: Active pretraining with successor features,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.028714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.633081Z digest=sha256:a5866f417ee3474e91bd8ff73f8d1bd1ba4d6354a3779f6b731a6e0a64de9d89

Observation da26a86e-3728-495a-9390-35aa84fa09de · outbound

This paper cites Efficient exploration via state marginal matching,.

Unsupervised Skill Discovery through Skill Regions Differentiation Efficient exploration via state marginal matching,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.017875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.636865Z digest=sha256:25d4da485bb1cac37b56498558ff1164663baaaf4249d45c95a376848a929859

Observation f32b9b01-64de-403f-95ca-bded11c4e539 · outbound

This paper cites Learning more skills through optimistic exploration,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning more skills through optimistic exploration,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.997823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.644518Z digest=sha256:d10f571115a8d3068b79049b0dc6f02f3bb58f55f3beaa983594c98a8352b796

Observation f6728e22-9bc1-4fde-bf6b-b38aeb7858cf · outbound

This paper cites Choreographer: Learning and adapting skills in imagination,.

Unsupervised Skill Discovery through Skill Regions Differentiation Choreographer: Learning and adapting skills in imagination,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.987631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.647901Z digest=sha256:1de678707c461adf1560b5d4267d048fceabc7002194a7d482b9b79b8c7914ba

Observation 57370c5e-b6cf-4ac9-a791-2a0f30a6ab51 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction,.

Unsupervised Skill Discovery through Skill Regions Differentiation Curiosity-driven exploration by self-supervised prediction,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.978763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.651425Z digest=sha256:ad8d13a7f2ebc758e239842ec453186d4d1e2f6eebe94a1508db61ceb10a759b

Observation cd209605-fe50-4286-bb76-0bd9626397f5 · outbound

This paper cites Self-supervised exploration via disagreement,.

Unsupervised Skill Discovery through Skill Regions Differentiation Self-supervised exploration via disagreement,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.968809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.655593Z digest=sha256:abd1de39acc16e3d4d2ce1caec28ec6953686f224f31321b2b973bbb0f74b672

Observation 6f595c2b-5c8a-4abe-95f3-e0ea93083b26 · outbound

This paper cites Exploration by random network distillation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Exploration by random network distillation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.959226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.659560Z digest=sha256:9e7d14fdbcc6423c39d38e551da89a4c04f802a807030e06b4221a8cdf06cceb

Observation e160b572-259a-4e80-b456-0ce9ac71b60d · outbound

This paper cites Reinforcement learn- ing with prototypical representations,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reinforcement learn- ing with prototypical representations,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.949840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.662915Z digest=sha256:72689ce0ff1bca09c2390de54dd744eb906f5e9b2df02e30ab9a3dcbbf1ef544

Observation 60d6b8f5-1755-43fa-be8b-c905890209f9 · outbound

This paper cites Unsupervised Skill-Discovery and Skill-Learning in Minecraft.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised Skill-Discovery and Skill-Learning in Minecraft

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.666160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.666160Z digest=sha256:6281ba212f68302419ca5b582aca14047dc222588807e24d61f5e9a6bded41a0

Observation 0174ae71-1d5d-4b66-b325-c726bebd3b3a · outbound

This paper cites Rethinking mutual information for language conditioned skill discovery on imitation learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Rethinking mutual information for language conditioned skill discovery on imitation learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.940307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.670115Z digest=sha256:e764df32c8a550accb851554cb5d1d16e061b883f0754eb6d8ed680395130602

Observation 87d5dae8-0c73-47fd-adea-1a04ce1e5fd5 · outbound

This paper cites Robust policy learning via offline skill diffusion,.

Unsupervised Skill Discovery through Skill Regions Differentiation Robust policy learning via offline skill diffusion,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.930800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.673501Z digest=sha256:295d5f8104e47b61053eab25db78ae872e4702194e67e220bad83a3f6a5b8cc7

Observation 532129b3-6699-4fdc-a237-47c38bbbd6c6 · outbound

This paper cites EUCLID: Towards efficient unsupervised reinforcement learning with multi-choice dynamics model,.

Unsupervised Skill Discovery through Skill Regions Differentiation EUCLID: Towards efficient unsupervised reinforcement learning with multi-choice dynamics model,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.920765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.676484Z digest=sha256:1f21da208805452c7da3c18db8d89d3a224d92325fdd8b52f68be70a8c5dd6ca

Observation 9b7851eb-dfa8-4001-93cb-37191848cf60 · outbound

This paper cites Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning.

Unsupervised Skill Discovery through Skill Regions Differentiation Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.680044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.680044Z digest=sha256:0adcea7631d349f4dcc677e20127c0dee503a09e336fc91acf040ac61b0e1988

Observation c0d40dd6-4f94-410f-9d1e-31652ed883fa · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice,.

Unsupervised Skill Discovery through Skill Regions Differentiation Deep reinforcement learning at the edge of the statistical precipice,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.910096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.683746Z digest=sha256:df963e8ae6988172739353efe06db015d6032a49b56b6fa0332458772584a380

Observation 66b480bf-5b9a-445a-93bd-0ddbd0d3441a · outbound

This paper cites DeepMind Control Suite.

Unsupervised Skill Discovery through Skill Regions Differentiation DeepMind Control Suite

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.688031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.688031Z digest=sha256:bd7f79e5a6ae3bd0b0d9ba6546202c2fea03c8e75eb8e8242b520b110e9c0576

Observation a9f30ed3-f669-4d5f-b7d8-e7b1c2a40bed · outbound

This paper cites Continuous control with deep reinforcement learning.

Unsupervised Skill Discovery through Skill Regions Differentiation Continuous control with deep reinforcement learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.691720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.691720Z digest=sha256:960b05c3bfb706c3390691fd3babc1e2b07308a3216f729fd01ccb662da2db5e

Observation 938e136d-270b-44d4-879d-efd40bcb47a9 · outbound

This paper cites Mastering visual continuous control: Improved data-augmented reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering visual continuous control: Improved data-augmented reinforcement learning,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.900348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.695397Z digest=sha256:c555f4d165a6bb6eff15283678b4baf2b99923400d748c70967694be5d8e47f9

Observation c14da304-2640-44aa-a0a0-67dbafd95f84 · outbound

This paper cites Mastering atari with discrete world models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering atari with discrete world models,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.890594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.699020Z digest=sha256:8896ceb7b6978cb2e54c8db0332d15b4469b861038f9f04e8c3184085251f49d

Observation 05d9e23b-2958-4fc4-b1d1-809d2a638e31 · outbound

This paper cites Provably efficient rein- forcement learning with linear function approximation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Provably efficient rein- forcement learning with linear function approximation,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.880897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.702736Z digest=sha256:3d76f097d422c75f83172b0ce55759a2fabccdd67c000d7b384588f293ebd7f0

Observation cf051aba-497b-4290-b167-116d05c26cdd · outbound

This paper cites Reward-free model-based reinforcement learning with linear function approximation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reward-free model-based reinforcement learning with linear function approximation,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.871448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.707318Z digest=sha256:b3eeb2eb15342b968058dea5491a2106cd73bd46d25f1daaa07519ac16e72388

Observation ae0bbdbe-c31f-4a3d-8441-e11393c504ed · outbound

This paper cites Deep variational information bottleneck,.

Unsupervised Skill Discovery through Skill Regions Differentiation Deep variational information bottleneck,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.711258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.711258Z digest=sha256:e6bb1e9ec63bb629602d27197dc7991fe41dd3b6d96a9172ecff246f4d0a01a6

Observation 66625992-49c9-4b03-abe3-5d4fee27eec8 · outbound

This paper cites Dynamic bottleneck for robust self-supervised exploration,.

Unsupervised Skill Discovery through Skill Regions Differentiation Dynamic bottleneck for robust self-supervised exploration,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.854381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.715141Z digest=sha256:8f8a86a1f7dc81927b7a3e5bb35abaa94648f58acf09190190ec75a49cc3dd3b

Observation 17e3e0cf-fcfd-4ab5-a163-7e6f6baa304e · outbound

This paper cites Provably efficient exploration in policy optimization,.

Unsupervised Skill Discovery through Skill Regions Differentiation Provably efficient exploration in policy optimization,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.844208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.718687Z digest=sha256:4f89b644d18d28d897efafab57d664f28cea9710e719855ad156206fccdd6242

Observation d95da145-63a3-467a-88a2-5c8a91764329 · outbound

This paper cites Logarithmic online regret bounds for undis- counted reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Logarithmic online regret bounds for undis- counted reinforcement learning,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.832856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.722538Z digest=sha256:5805ccbac5c8336c977d25922ff6e4fe5e96f1563d68a002c3bdfdd8c9acb359

Observation 591aa826-3252-4443-a02e-6acff41e581c · outbound

This paper cites Available: https://openreview.net/forum?id=Hkla1eHFvS.

Unsupervised Skill Discovery through Skill Regions Differentiation Available: https://openreview.net/forum?id=Hkla1eHFvS

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.008510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:27:00.640914Z digest=sha256:8a470805cd9ea28f99bdc5fa1bd16c5c6ad05fc575af75e17fa08737de2e9ad1

Pith citing papers

No inbound Pith citation observations are available.