Pith. sign in

Paper Citation Record · LEDGER

Dynamics-Aware Unsupervised Discovery of Skills

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:1907.01657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.01657 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:11:08.934687Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:39:34.298932Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5acb6b12-c7c7-4c9e-accb-c3b05f64da49 · inbound

Is Conditional Generative Modeling all you need for Decision-Making? cites this paper.

Is Conditional Generative Modeling all you need for Decision-Making? Dynamics-Aware Unsupervised Discovery of Skills

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:35:10.807333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T15:35:10.593969Z digest=sha256:27f3cc7e21055873d19a11d1ff957541adf6d2f75958ebd2ea929bebe06fc78a

Observation 7f974e1c-aa56-4a94-9545-65687568ef72 · inbound

Training RL Agents for Multi-Objective Network Defense Tasks cites this paper.

Training RL Agents for Multi-Objective Network Defense Tasks Dynamics-Aware Unsupervised Discovery of Skills

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:11:08.934687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:11:08.934687Z digest=sha256:56bca01796a02180a774847ebe94d1b4738a5f54741df84f2df5d8938f195741

Observation e03d3241-ec77-42e9-b744-9191bd314146 · inbound

RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming cites this paper.

RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming Dynamics-Aware Unsupervised Discovery of Skills

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:20.279140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:20.279140Z digest=sha256:079cba6030ee38a87e6a99b5c2c54113900726cd19021f9844c2d61ea0abc897

Observation 6c34781e-890b-4cc6-b75c-9f2be1a45cb6 · inbound

Epistemically-guided forward-backward exploration cites this paper.

Epistemically-guided forward-backward exploration Dynamics-Aware Unsupervised Discovery of Skills

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:32:34.807117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:32:34.807117Z digest=sha256:f448324b59c458c078ee672b6293c2a99461d734db13fbb2d976c6e0c20d9186

Observation 15e5ecc8-53f5-4941-bf8a-32a8a5e781e0 · inbound

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning cites this paper.

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T15:56:59.109038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:56:59.109038Z digest=sha256:f109a8400d581d4d04d0528fd8eab06a45e3f57376a3cbe38e225ca617820622

Observation 278294fc-c0de-4947-8e82-edd4710c55d0 · inbound

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs cites this paper.

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs Dynamics-Aware Unsupervised Discovery of Skills

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:34.996582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:17:34.996582Z digest=sha256:1b231498e2cf8aaee3aa8c667ea0a975b5a1e4f045a9cef0b49ccf51f5e95831

Observation f91f9964-d241-4916-963c-270cb71cd320 · inbound

Hierarchical Behaviour Spaces cites this paper.

Hierarchical Behaviour Spaces Dynamics-Aware Unsupervised Discovery of Skills

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.758712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T03:33:48.527621Z digest=sha256:c7e1a11ff3d42a41a7cafe3f0c7b5cceed7747b5cb07a41007e8fc9713613486

Observation f3455a21-b401-4303-8cf9-5db592cdf331 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Dynamics-Aware Unsupervised Discovery of Skills

Reference 165

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:58.267267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:cfc28663363c48777156dce8bd6c0b41abfa803e5bec675c8ef952f63094cc54

Observation e924b165-9681-4871-ab40-5a8760a8dc21 · inbound

Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization cites this paper.

Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization Dynamics-Aware Unsupervised Discovery of Skills

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:46:10.749389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T13:52:43.033237Z digest=sha256:cece59d83e7e073ffae051e62d759bb4af9fdc60a114a03cafc2d30eacb8e0fc

Observation 686c3abf-358b-43e8-93d9-8e92095cf7e7 · inbound

Delay-Empowered Causal Hierarchical Reinforcement Learning cites this paper.

Delay-Empowered Causal Hierarchical Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:47:21.205769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T05:46:51.659283Z digest=sha256:de9c86cf966c29ce37de8f2fd8dc66f7a1e9c0925a1dc1da4e25a376c87bdfec

Observation 1f4140f7-9c74-463d-bec6-c24b759d7fe7 · inbound

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning cites this paper.

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:26.372151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T18:37:48.259353Z digest=sha256:db4275a085f9b0eaf8a30333fc9a0dd048e07eeceac8591bdafe4dcc57b9f841

Observation 03c1cede-f8a5-4dd2-8a67-1c86337e9095 · inbound

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling cites this paper.

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling Dynamics-Aware Unsupervised Discovery of Skills

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:07:34.177073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T16:07:28.361043Z digest=sha256:57aced0072ca8d4831c3c57fec74b1596ffdcf73cb739f661af9a54a2a01c9d3

Observation 10b55895-2a64-4cd9-b449-f9a9328fd0fc · inbound

Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory cites this paper.

Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory Dynamics-Aware Unsupervised Discovery of Skills

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:39:34.301049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T16:53:53.441274Z digest=sha256:932d7d6f957c5052fa1e15211034c9690ea01777eb42a2890cb50fe880b6b983

Observation b7d65f84-7599-455f-a48e-20e67a2266d8 · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering Dynamics-Aware Unsupervised Discovery of Skills

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.206134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:593f452706aafb6d9dda0f1c5a1603e12af498b234a557e3a73c3255e1b5f8c5

Observation 534371e1-d8ba-4b70-9bdd-0db6848e22e7 · inbound

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies cites this paper.

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies Dynamics-Aware Unsupervised Discovery of Skills

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T06:32:52.458556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:32:52.458556Z digest=sha256:2854aef688c50dbb801ee214a14527dfad8b2513b0197c318a740138ce56de98

Observation 1d53a89b-01a5-44af-929d-7b3fbc749eb3 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Dynamics-Aware Unsupervised Discovery of Skills

Reference 229

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:35.216572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:35.216572Z digest=sha256:2c3e68d3fdffc3a39dae24bf13aec00ad779207dc3d0845cc2ff127c60df6121