Pith. sign in

Paper Citation Record · LEDGER

Dynamics-Aware Unsupervised Discovery of Skills

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:1907.01657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.01657 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:11:08.934687Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:39:34.298932Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5acb6b12-c7c7-4c9e-accb-c3b05f64da49 · inbound

Is Conditional Generative Modeling all you need for Decision-Making? cites this paper.

Is Conditional Generative Modeling all you need for Decision-Making? Dynamics-Aware Unsupervised Discovery of Skills

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:35:10.807333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T15:35:10.593969Z digest=sha256:0cef8895be364428e5e071797e3d56382c24b66155cf8b0e8792b152a3317e5b

Observation 7f974e1c-aa56-4a94-9545-65687568ef72 · inbound

Training RL Agents for Multi-Objective Network Defense Tasks cites this paper.

Training RL Agents for Multi-Objective Network Defense Tasks Dynamics-Aware Unsupervised Discovery of Skills

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:11:08.934687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:11:08.934687Z digest=sha256:c87846f9c61af6d08a38851e54911d932d80a8b034938fc61f31f411f9809946

Observation e03d3241-ec77-42e9-b744-9191bd314146 · inbound

RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming cites this paper.

RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming Dynamics-Aware Unsupervised Discovery of Skills

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:20.279140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:20.279140Z digest=sha256:079cba6030ee38a87e6a99b5c2c54113900726cd19021f9844c2d61ea0abc897

Observation 6c34781e-890b-4cc6-b75c-9f2be1a45cb6 · inbound

Epistemically-guided forward-backward exploration cites this paper.

Epistemically-guided forward-backward exploration Dynamics-Aware Unsupervised Discovery of Skills

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:32:34.807117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:32:34.807117Z digest=sha256:f448324b59c458c078ee672b6293c2a99461d734db13fbb2d976c6e0c20d9186

Observation 15e5ecc8-53f5-4941-bf8a-32a8a5e781e0 · inbound

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning cites this paper.

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T15:56:59.109038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:56:59.109038Z digest=sha256:f109a8400d581d4d04d0528fd8eab06a45e3f57376a3cbe38e225ca617820622

Observation 278294fc-c0de-4947-8e82-edd4710c55d0 · inbound

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs cites this paper.

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs Dynamics-Aware Unsupervised Discovery of Skills

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:34.996582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:17:34.996582Z digest=sha256:1b231498e2cf8aaee3aa8c667ea0a975b5a1e4f045a9cef0b49ccf51f5e95831

Observation f91f9964-d241-4916-963c-270cb71cd320 · inbound

Hierarchical Behaviour Spaces cites this paper.

Hierarchical Behaviour Spaces Dynamics-Aware Unsupervised Discovery of Skills

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.758712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T03:33:48.527621Z digest=sha256:55d1ed01027b739c98daa7341430edccfe96f8c2640e1b3811ce2b05003a3b5c

Observation f3455a21-b401-4303-8cf9-5db592cdf331 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Dynamics-Aware Unsupervised Discovery of Skills

Reference 165

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:58.267267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:9a07b9ad2007fee9a960b3368e1604cbeea39f9f93814f1f7d62a4511d29ffb9

Observation e924b165-9681-4871-ab40-5a8760a8dc21 · inbound

Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization cites this paper.

Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization Dynamics-Aware Unsupervised Discovery of Skills

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:46:10.749389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T13:52:43.033237Z digest=sha256:783d74f981b3267586e54fb0624c076ce82a50c2f7bd8639cae753a09feaa14b

Observation 686c3abf-358b-43e8-93d9-8e92095cf7e7 · inbound

Delay-Empowered Causal Hierarchical Reinforcement Learning cites this paper.

Delay-Empowered Causal Hierarchical Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:47:21.205769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T05:46:51.659283Z digest=sha256:9e0a5947fdef774d541454334c37911b4772433f6f82a6d51ac898556a87deed

Observation 1f4140f7-9c74-463d-bec6-c24b759d7fe7 · inbound

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning cites this paper.

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:26.372151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T18:37:48.259353Z digest=sha256:ab47f3e21d906c93fd8f30a930c06a0a8f3d7ce7c0455e60df3747e96d2f9bd8

Observation 03c1cede-f8a5-4dd2-8a67-1c86337e9095 · inbound

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling cites this paper.

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling Dynamics-Aware Unsupervised Discovery of Skills

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:07:34.177073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T16:07:28.361043Z digest=sha256:5e37b633cc7995d1634f07070442ac4da0097793d01e0b244e8400ea256b7c44

Observation 10b55895-2a64-4cd9-b449-f9a9328fd0fc · inbound

Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory cites this paper.

Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory Dynamics-Aware Unsupervised Discovery of Skills

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:39:34.301049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T16:53:53.441274Z digest=sha256:0ee065423c7ebef5f14e8f278cb18a80863db411682f75f1ba922c415343830c

Observation b7d65f84-7599-455f-a48e-20e67a2266d8 · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering Dynamics-Aware Unsupervised Discovery of Skills

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.206134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:890fa93297fff8be2b5dc35b06ee5ebc884e625f86b07c23eec2e7b005dc3276

Observation 534371e1-d8ba-4b70-9bdd-0db6848e22e7 · inbound

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies cites this paper.

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies Dynamics-Aware Unsupervised Discovery of Skills

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T06:32:52.458556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:32:52.458556Z digest=sha256:2854aef688c50dbb801ee214a14527dfad8b2513b0197c318a740138ce56de98

Observation 1d53a89b-01a5-44af-929d-7b3fbc749eb3 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Dynamics-Aware Unsupervised Discovery of Skills

Reference 229

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:35.216572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:35.216572Z digest=sha256:2c3e68d3fdffc3a39dae24bf13aec00ad779207dc3d0845cc2ff127c60df6121