Pith. sign in

Paper Citation Record · LEDGER

Efficient Skill Discovery via Regret-Aware Optimization

As of 12 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 0 inbound Pith citation observations for arXiv:2506.21044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21044 v1

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:41:38.230312Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

74 of 74 outbound references displayed

  • verified exact2
  • verified fuzzy51
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 90c85963-0971-41b0-9aa5-55b627618389 · outbound

This paper cites write newline.

Efficient Skill Discovery via Regret-Aware Optimization write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.145979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.145979Z digest=sha256:3e7f05682d5322be76cc30f3b1e6bb0e401774fd501e9a9122db8f7e4f1cb207

Observation 1da8fa44-ed1a-4839-883e-e382ad82a665 · outbound

This paper cites M., Crump, T., and Far, B.

Efficient Skill Discovery via Regret-Aware Optimization M., Crump, T., and Far, B

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.322974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.201982Z digest=sha256:a856e40e30fe3676cb22e466acd34fedd1dfe5932c094daaf5780720ee8b1a04

Observation b46a063f-888c-427f-b844-28504c636e07 · outbound

This paper cites Hindsight Experience Replay.

Efficient Skill Discovery via Regret-Aware Optimization Hindsight Experience Replay

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.306850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.306850Z digest=sha256:d4011d874e449c20683495c54bcc812ff0145dc66c9d753c7af8a43fd3031c3d

Observation 88f1255f-08d8-45f4-a571-a72faee19108 · outbound

This paper cites TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations.

Efficient Skill Discovery via Regret-Aware Optimization TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.424348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.424348Z digest=sha256:11be66b032e5903e540f73dafe9f9cfaa40b9be2bccfcd13c980443065cbe0f6

Observation 1ea52e5c-7a75-40e5-8676-473c33a22ed3 · outbound

This paper cites K., and Konidaris, G.

Efficient Skill Discovery via Regret-Aware Optimization K., and Konidaris, G

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.310533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.518360Z digest=sha256:a2486844f6ea0eff22455a8e7c8f250b6364c9ec91b46eae034a71a2265ee986

Observation 0e414a57-7eaf-4db7-9a17-26479fb26d17 · outbound

This paper cites Constrained ensemble exploration for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Constrained ensemble exploration for unsupervised skill discovery

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.297553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.609290Z digest=sha256:3db7cb3bf65213e4f5de459bb4d4dbb0ce29676d3a9b4b6a661b78bc89fb921d

Observation 9a4094e2-9f68-4cd4-bdfe-ea633f5bcc5d · outbound

This paper cites X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U.

Efficient Skill Discovery via Regret-Aware Optimization X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.283764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.703614Z digest=sha256:b622ac7b1f52470c9171b0a4dffa7cdd85f9249d8cd9d7db84bec5b92157782f

Observation 0f826a26-1c03-4767-ade4-ceeb660bfa10 · outbound

This paper cites Exploration by random network distillation.

Efficient Skill Discovery via Regret-Aware Optimization Exploration by random network distillation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.270252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.785283Z digest=sha256:fcad4156b2e6061de8d41acbd92b23ac2ead0d291278e57fb88cfe7d2b38fac4

Observation 55067e30-d11b-4a7b-9452-07042e042327 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills.

Efficient Skill Discovery via Regret-Aware Optimization Explore, discover and learn: Unsupervised discovery of state-covering skills

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.256300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.836290Z digest=sha256:95da1289dd06844370d01823f2e41ce788aea03e2cc38a05388e54eca3e298f1

Observation 6f836df7-b49c-4975-87b2-7d70de188e82 · outbound

This paper cites Language as a cognitive tool to imagine goals in curiosity driven exploration.

Efficient Skill Discovery via Regret-Aware Optimization Language as a cognitive tool to imagine goals in curiosity driven exploration

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.241660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.917944Z digest=sha256:7aeb0588025c54520af0fa754ab77da434b5b6580f34ac22fcce71c7bc18f874

Observation 0b03d062-5f28-41ad-bee5-78a4176256fd · outbound

This paper cites Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey.

Efficient Skill Discovery via Regret-Aware Optimization Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.227451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.005894Z digest=sha256:8c194d0d8604a913f3d998cc344cc73e3521656d571dfe93b31c6dc586b26e3e

Observation ba3fe812-2a9d-4457-abdc-473cae2bb186 · outbound

This paper cites Goal-conditioned imitation learning.

Efficient Skill Discovery via Regret-Aware Optimization Goal-conditioned imitation learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.212302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.090668Z digest=sha256:ecc4da601aa0f1f08f50250df165b1ca76343d44dc10a4de87b66ba655c431cb

Observation a5494acd-cb81-4cd8-a6e5-f246fab0f8bd · outbound

This paper cites Adversarial intrinsic motivation for reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Adversarial intrinsic motivation for reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.198143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.203128Z digest=sha256:f1cf71fe89d8c2115cadd59893d09e961dfe1bef6b3e2ef9c85a31ee5ab047dd

Observation e284d72f-c053-4426-b2b3-0abe976a1ce9 · outbound

This paper cites Diversity is all you need: Learning skills without a reward function.

Efficient Skill Discovery via Regret-Aware Optimization Diversity is all you need: Learning skills without a reward function

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.184132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.247740Z digest=sha256:2f35ddab9e06a8681595ff62c06a34cd278c48359a77fa624a7c23ebe93b5401

Observation 143187d5-3f4d-4258-9ef2-695f91139aa3 · outbound

This paper cites C-Learning: Learning to Achieve Goals via Recursive Classification.

Efficient Skill Discovery via Regret-Aware Optimization C-Learning: Learning to Achieve Goals via Recursive Classification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.332286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.332286Z digest=sha256:199acd485f5380a42ea34d403bdda18b04ee08d2db0a2bdb0f474fb2458fbde8

Observation 20c0ea5b-129c-4409-bda9-7de5ec4a561b · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.170092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.417950Z digest=sha256:10a10a1108fb5a1ddbc6c5edc631757ff208c05d6e759fd2b8210c6ad52f8691

Observation d9147cd4-2c24-4fa6-b982-d8eec0274326 · outbound

This paper cites Curriculum-guided hindsight experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Curriculum-guided hindsight experience replay

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.156692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.468891Z digest=sha256:e7312c5f1398667f9bc1908da5664e11a030205382dafcf909f9a7a12cbc336d

Observation a5e18089-2bb6-4ebf-b2dd-aa137296a80a · outbound

This paper cites Automatic goal generation for reinforcement learning agents.

Efficient Skill Discovery via Regret-Aware Optimization Automatic goal generation for reinforcement learning agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.473619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.473619Z digest=sha256:bca5f22d309ab0679c292a1364c8e5dbc51cc451d7b4e5721bc580ba5d95b6ed

Observation 7d6cf578-631c-4d71-849a-32e0cf78ae83 · outbound

This paper cites Accuracy-based Curriculum Learning in Deep Reinforcement Learning.

Efficient Skill Discovery via Regret-Aware Optimization Accuracy-based Curriculum Learning in Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.578090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.578090Z digest=sha256:da3cde4b4dec9cee7bb0ad0902f49e9eb56ea3f618be14f7334420f51fab127f

Observation ccb63044-06ba-4f5f-8633-324719177297 · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2020.

Efficient Skill Discovery via Regret-Aware Optimization D4rl: Datasets for deep data-driven reinforcement learning, 2020

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.133429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.762057Z digest=sha256:9b3c8f14cbb34ebaa09c4e5c3556c89b21d86502fb08a25c5f131262c395924c

Observation 77335ca6-e087-4292-8657-b2892f585840 · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Efficient Skill Discovery via Regret-Aware Optimization Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.963486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.963486Z digest=sha256:c0ef7d3789ebfe1ee15523f76fb3fb50ae328e82d2fd419bb2f895d744e3514d

Observation 53a1012f-57ab-4f05-a980-b04548088400 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.121608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.121608Z digest=sha256:f348a92bd8558ed8d0e3964b7070b6a25a2c37bc76c6c5a9ce960b6a2fbdfe85

Observation 50abf721-f062-4549-8f28-ca8c2852b4eb · outbound

This paper cites J., and Wierstra, D.

Efficient Skill Discovery via Regret-Aware Optimization J., and Wierstra, D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.110596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.261687Z digest=sha256:3669c351ca2f0845d02918a22a1c9b5e24757b070042776619b2483d18870bc9

Observation 37668b98-4ab1-4136-be9f-db203f06ca2b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Efficient Skill Discovery via Regret-Aware Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.471301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.471301Z digest=sha256:12da9976a6e9d20f5ac6bb2bfdf7d483ebe092e6bcb740f3903c201b9fd829a5

Observation a1cbfc2f-458d-472e-a9bb-20c95be91d63 · outbound

This paper cites Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation.

Efficient Skill Discovery via Regret-Aware Optimization Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.461935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.692807Z digest=sha256:38da2899c5f1d3b29e72639a30298ee3d2e88e9b17040e7946697493ae166c8b

Observation 33de2148-99d3-48aa-8870-312b786ce215 · outbound

This paper cites Exploration in deep reinforcement learning: From single-agent to multiagent domain.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: From single-agent to multiagent domain

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.089016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.858079Z digest=sha256:e47d9d376016b477fd156ff41b635f96c3202108728a9db455727da70ad45e19

Observation 3a774f42-b30f-4306-8538-3989effc3200 · outbound

This paper cites Open-Endedness is Essential for Artificial Superhuman Intelligence.

Efficient Skill Discovery via Regret-Aware Optimization Open-Endedness is Essential for Artificial Superhuman Intelligence

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.029860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.029860Z digest=sha256:485b6dc4eccada4093ac4e8cd11d901393bf2836b08c82485371e2ae2050c0f2

Observation ce161550-23c4-49d2-a496-0089d72ad128 · outbound

This paper cites Unsupervised curricula for visual meta-reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised curricula for visual meta-reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.074905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.228723Z digest=sha256:fc62e2103edbe60f2121f0c7b8afae3317f1884dbe25c33a0ab458481d6d194e

Observation 63d64998-cf2b-468f-9fe6-c4c89b7cb46e · outbound

This paper cites A comprehensive survey on self-interpretable neural networks.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey on self-interpretable neural networks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.398189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.398189Z digest=sha256:e9fb448cb2cf8bb810961c8ad6acb4b9f1f2beecb8675995d1a5b79eb27c6376

Observation f843fe81-8fb2-4a69-a52a-25fa443fe92a · outbound

This paper cites Replay-Guided Adversarial Environment Design.

Efficient Skill Discovery via Regret-Aware Optimization Replay-Guided Adversarial Environment Design

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:41:38.328725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.601070Z digest=sha256:375cbb505a0267501b386653337b6cac20c6f351b91d9f0f0c14a7c4256b2be4

Observation 1fabd924-f58b-4583-b50b-12a9223ad4f0 · outbound

This paper cites Prioritized level replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized level replay

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.061604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.669110Z digest=sha256:8ce8ceaf1032453f0051db4f37fe00d626b6cee6e8334aa7150a3abeb0d24793

Observation d4f15ff1-9b76-4682-98e1-2f2825d64cbe · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.047719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.714340Z digest=sha256:1a1eb446fa2da63100461e0f142aa877b3006eb5c68336cef734a829bcc5459c

Observation 844b097e-fc4a-44d9-959f-3033e3aa5941 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.740268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.740268Z digest=sha256:c03b4158922cf77ecece9b55088a181d74b913cabf50e0a8591e211a58a7d7c3

Observation 3b147d4b-9840-417d-a8ce-3fcd0f63d7b5 · outbound

This paper cites K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J.

Efficient Skill Discovery via Regret-Aware Optimization K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.026333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.814734Z digest=sha256:5106d99e693b55c9fd1890e8d74e6b1e20d6d56169724fb484e559d93b3fcd77

Observation e96c04cd-b09c-4589-87e6-e8d3b65bec21 · outbound

This paper cites Unsupervised skill discovery with bottleneck option learning, 2021.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised skill discovery with bottleneck option learning, 2021

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.012568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.865430Z digest=sha256:942fcadac3585559c89f02a62be05c82b8c55f41e6ded8c72786f38c6434db16

Observation 31c84aeb-d9aa-468b-842f-3722166e7392 · outbound

This paper cites Active world model learning with progress curiosity.

Efficient Skill Discovery via Regret-Aware Optimization Active world model learning with progress curiosity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.998539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.942087Z digest=sha256:7274881a7cb219bb50690e3310c1166279a473ffd60f3908e25e07cc8a24e147

Observation 661f3f1b-5f3b-4fb8-947c-aef0e73963a1 · outbound

This paper cites Exploration in deep reinforcement learning: A survey.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: A survey

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.985466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.015977Z digest=sha256:a95ff1a92485007bd41d10d96529a740822bb50ebd3bb889eb7a8072db475567

Observation b6d116a6-a48f-4cab-8144-738129961de2 · outbound

This paper cites B., Yarats, D., Rajeswaran, A., and Abbeel, P.

Efficient Skill Discovery via Regret-Aware Optimization B., Yarats, D., Rajeswaran, A., and Abbeel, P

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.971539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.079101Z digest=sha256:f6eea89dc3cca51810a11fada34477784013142b02ab6835d710f59d1613a227

Observation 7fb3eaea-27e2-4d7f-b4c3-a3b8e91864d3 · outbound

This paper cites and Seo, S.-W.

Efficient Skill Discovery via Regret-Aware Optimization and Seo, S.-W

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.957700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.143371Z digest=sha256:7b8c412e25a2ccb07ce853de8eb0669a8beef2fe8ed74a028feffbc6d9711a05

Observation 95fd50be-07ad-4499-b263-8839f0ba694b · outbound

This paper cites Revisiting graph adversarial attack and defense from a data distribution perspective.

Efficient Skill Discovery via Regret-Aware Optimization Revisiting graph adversarial attack and defense from a data distribution perspective

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.944068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.200699Z digest=sha256:514180da15cc5cabbff3ea73dbbc6966c3c1970ac1393975ea2750c4a3d8e486

Observation 5cbd798b-7508-4414-98dd-1f02158d4e1f · outbound

This paper cites Boosting the adversarial robustness of graph neural networks: An ood perspective.

Efficient Skill Discovery via Regret-Aware Optimization Boosting the adversarial robustness of graph neural networks: An ood perspective

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.929948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.273877Z digest=sha256:27112ac3e20d4278bd36b0d1b430696b0ee447f1b6fb1e33d13182299e46f458

Observation 8972b79e-4f56-4c4d-879d-e3ff63f6d453 · outbound

This paper cites A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals.

Efficient Skill Discovery via Regret-Aware Optimization A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:37.321403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:37.321403Z digest=sha256:f561f001efb857525db937fe40024a0a788461e0fe0b9aa5dd987428d8d6899a

Observation 17c92cf3-8124-4217-8929-958f476c3f8b · outbound

This paper cites Choreographer: Learning and adapting skills in imagination.

Efficient Skill Discovery via Regret-Aware Optimization Choreographer: Learning and adapting skills in imagination

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.916681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.365884Z digest=sha256:a15daf462f25aaba5c7107048bb8362346c96e5e4bb16d5f9123e9cddfe4a0f8

Observation 57d2ebd6-1b94-4784-b061-afed5604589b · outbound

This paper cites Planning with goal-conditioned policies.

Efficient Skill Discovery via Regret-Aware Optimization Planning with goal-conditioned policies

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.901968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.462003Z digest=sha256:9b54d67ca2f77737457f20ff8e6cfde0e094069ff0e2a410b0c0911c3d04cfb0

Observation 06538529-6818-4526-94fe-164caa8d0a93 · outbound

This paper cites Wasserstein dependency measure for representation learning.

Efficient Skill Discovery via Regret-Aware Optimization Wasserstein dependency measure for representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.889465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.499235Z digest=sha256:31cb1a0b33d5b8cacad54e6a96ad5a0d828f463644337bad14ea56480fc28ecb

Observation cb1c474e-9636-4702-886b-d7f66500ee3f · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Lipschitz-constrained unsupervised skill discovery

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.876381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.582284Z digest=sha256:eb82892279ab903e3c7013969a3a2a607f29314568448f85f6007d8c936b9754

Observation 2318ca52-4a84-4898-b339-788fcc50afb8 · outbound

This paper cites Hiql: Offline goal-conditioned rl with latent states as actions.

Efficient Skill Discovery via Regret-Aware Optimization Hiql: Offline goal-conditioned rl with latent states as actions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.863376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.652587Z digest=sha256:07acbebc0403789ba2f308cd7ab38b6a61a5ec68f7a062881cfd481a1ddf01c2

Observation edf572e2-741e-428e-8fa2-e7cf5223703b · outbound

This paper cites Foundation policies with hilbert representations.

Efficient Skill Discovery via Regret-Aware Optimization Foundation policies with hilbert representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.849816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.697677Z digest=sha256:17014d745add12111665f8d06306717e50dfc36262a54ac49f32d28d8b928a3a

Observation 1c666100-dff7-49e6-9bb4-4666000b500c · outbound

This paper cites METRA : Scalable unsupervised RL with metric-aware abstraction.

Efficient Skill Discovery via Regret-Aware Optimization METRA : Scalable unsupervised RL with metric-aware abstraction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.837059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.797162Z digest=sha256:7638c9d264efa936ceb0c75437db94e2f1cddfd13afff89654e932de63bf428c

Observation b04ca4b0-c24e-47a8-a6e3-beb4780880fa · outbound

This paper cites Evolving curricula with regret-based environment design.

Efficient Skill Discovery via Regret-Aware Optimization Evolving curricula with regret-based environment design

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.824248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.867902Z digest=sha256:dfd96885f79563ad873c819cf94d2d68e850f42f75b05abe293a5b5872e6f560

Observation 06fc72e3-eb0c-41b2-aeec-450f54d74c22 · outbound

This paper cites A., and Darrell, T.

Efficient Skill Discovery via Regret-Aware Optimization A., and Darrell, T

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.810397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.943248Z digest=sha256:ec511011b8d2e586d3238e745d442f8fe8b8bf8b4d6896daf083de714245cbd9

Observation 99fba55d-032c-4e6b-9bbb-6e0f2bb19b85 · outbound

This paper cites H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S.

Efficient Skill Discovery via Regret-Aware Optimization H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.796256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.036854Z digest=sha256:ba4603313767f48d662f249b9e470970ce6c12bf2607690f36906ceb3de6b8b6

Observation 9cfa1ce2-4d30-42d0-8d0b-d85150d118b1 · outbound

This paper cites Automatic Curriculum Learning For Deep RL: A Short Survey.

Efficient Skill Discovery via Regret-Aware Optimization Automatic Curriculum Learning For Deep RL: A Short Survey

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.079354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.079354Z digest=sha256:3ccefd7511f2b8b3a40bcc56b1b5065a9906a9a841cc0e2aa47c57a352dd5f50

Observation 428819a6-0947-4e3d-90bb-107c5f699bdb · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:38.783134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.121688Z digest=sha256:40d1fd87e629e4c50c9c418af6198065ed7861463a02e025754e646dcd6f8de8

Observation 20125f29-6832-4238-89d4-625fbd6c8abc · outbound

This paper cites Prioritized experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized experience replay

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.770421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.131699Z digest=sha256:aa19199638e57191657c76377986341b208d0f80caf182f69c2096fd5f67066a

Observation c49c44bb-ec1b-4384-93b5-5b9283d4a7b8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Proximal Policy Optimization Algorithms

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.139133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.139133Z digest=sha256:5c06486aa4309318cb9661eb08e78a3b6e4247f023337c7f4c36e6ba65762bbe

Observation 37147148-54ea-48b5-ae05-9d0211ef2774 · outbound

This paper cites Dynamics-aware unsupervised discovery of skills.

Efficient Skill Discovery via Regret-Aware Optimization Dynamics-aware unsupervised discovery of skills

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.755575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.149622Z digest=sha256:0182f9bfd2f58ad47a5aaf7f4fcaf2684ecde2c436ddf7516df54600c6876b6a

Observation 3ece75c2-8146-4a87-aac1-db7807dc5998 · outbound

This paper cites Deterministic policy gradient algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Deterministic policy gradient algorithms

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.740517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.157683Z digest=sha256:3e4fae9b8b3af3b029c3baf938d704047446318104e38f960ebc42cd576b8366

Observation ae196208-c010-45c4-9dee-eea8293b97e0 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play.

Efficient Skill Discovery via Regret-Aware Optimization A general reinforcement learning algorithm that masters chess, shogi, and go through self-play

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.725250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.165986Z digest=sha256:3cc7a1cfeb4c6c2fe35e6b02f3d255522087f3621563993237a202891051c505

Observation 05395d97-732b-444f-971b-fd92e2885772 · outbound

This paper cites Intrinsic motivation and automatic curricula via asymmetric self-play.

Efficient Skill Discovery via Regret-Aware Optimization Intrinsic motivation and automatic curricula via asymmetric self-play

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.710934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.170823Z digest=sha256:4e031175ef6dfad42b208a91fda76912016be14fdb78f3699a873a6a5ec1555e

Observation 72c02ece-e109-43a7-a90a-62125a6fec8f · outbound

This paper cites Policy continuation with hindsight inverse dynamics.

Efficient Skill Discovery via Regret-Aware Optimization Policy continuation with hindsight inverse dynamics

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.696597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.174661Z digest=sha256:39e919a2a2c52af641298699c3ee7ad7506529e0bcfda66a34d38c4548c6f514

Observation 9343b173-20df-48ed-9c25-1f9d65641445 · outbound

This paper cites Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections.

Efficient Skill Discovery via Regret-Aware Optimization Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.682455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.179067Z digest=sha256:33a6a185f6f53b6c05c99fb4bdca08595655dec095f437fb59d4fd1c38c24789

Observation fdad43a7-eb70-470d-baff-90cdd4c02170 · outbound

This paper cites Market-aware long-term job skill recommendation with explainable deep reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Market-aware long-term job skill recommendation with explainable deep reinforcement learning

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.667895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.183587Z digest=sha256:6b103c36c59a723e179becc0f7ba506f5edd4afd6de6ba780be497b34e992756

Observation 8abb6957-0f84-4e6a-b131-b2d2787a6ba8 · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.188231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.188231Z digest=sha256:f0ebf1d3f430fcf13cc856741afa026e69c572e4d10fb43e29db284b29757324

Observation 2c1ae144-c4ff-4eb7-bd6d-223d2bb13244 · outbound

This paper cites S., McAllester, D., Singh, S., and Mansour, Y.

Efficient Skill Discovery via Regret-Aware Optimization S., McAllester, D., Singh, S., and Mansour, Y

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.191955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.191955Z digest=sha256:c3bf7bbe7214510ebdd8886e708babc25bb101d02eb1c3f1a633f78a60ccad72

Observation 0c85fde9-bacb-4a1b-8a53-ba00562a2777 · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Efficient Skill Discovery via Regret-Aware Optimization dm\_control: Software and tasks for continuous control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.637023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.196846Z digest=sha256:105b167c424472a967ee4fb2a11f22cddfc169abea8f5050308ca71814467260

Observation 7fe00edc-7e18-485c-8903-9a3394f52bc8 · outbound

This paper cites M., Mathieu, M., Dudzik, A., Chung, J., Choi, D.

Efficient Skill Discovery via Regret-Aware Optimization M., Mathieu, M., Dudzik, A., Chung, J., Choi, D

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.623333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.201024Z digest=sha256:fe275cc5d3f79cb7c1c2e4a696f7bc633db2ccda5bea43051ef703183610d9ab

Observation d7b4d0c5-2d5f-4bdc-84e9-52b926ab9d66 · outbound

This paper cites Optimal goal-reaching reinforcement learning via quasimetric learning.

Efficient Skill Discovery via Regret-Aware Optimization Optimal goal-reaching reinforcement learning via quasimetric learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.608881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.204893Z digest=sha256:bd50728d16240413e67612419c911c97da3b3e6a4b66b939d4d38c111146fa0b

Observation 366bb4e3-00e8-4f0d-9c7f-15d14fc0b0d7 · outbound

This paper cites Ski LD : Unsupervised skill discovery guided by factor interactions.

Efficient Skill Discovery via Regret-Aware Optimization Ski LD : Unsupervised skill discovery guided by factor interactions

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.594975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.209112Z digest=sha256:abda71fde14825bf67a2990ceacd0814a08eecb46172ed2701fe2560369858cc

Observation 14c24df4-3b5b-43a6-8166-d494a170eb41 · outbound

This paper cites A comprehensive survey of forgetting in deep learning beyond continual learning.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey of forgetting in deep learning beyond continual learning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.581685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.213181Z digest=sha256:e31a5e418668b51ca94e1eadff8a2651360c87f354fb657335f985b7d4065070

Observation 2d352df8-1244-410b-94cb-5d6481ba4334 · outbound

This paper cites Neural Program Synthesis By Self-Learning.

Efficient Skill Discovery via Regret-Aware Optimization Neural Program Synthesis By Self-Learning

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.271036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.217400Z digest=sha256:cbd39b2a9214d12234d4f8d0d0670b5de059272f259f3ada6fe3b833051072f5

Observation 499e79d3-9c9e-4f05-ba46-034c0a95d1b5 · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Behavior contrastive learning for unsupervised skill discovery

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.568005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.222361Z digest=sha256:1b55b3cf6d7c534e79a9f2865d0efbc822977f011d213da500949d52599c5394

Observation 9db995b6-e3eb-4e5b-9d9b-30ce1be6506b · outbound

This paper cites Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.553086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.226237Z digest=sha256:75753acc2c2a27f4448570bc81bb791de8d2d9769047de0ced71f1ff5cbdd963

Observation 6c028a05-0532-41b4-8993-8abd49123dec · outbound

This paper cites Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach.

Efficient Skill Discovery via Regret-Aware Optimization Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.538713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.230312Z digest=sha256:1fcecc5f32a4ecd783da1f249be0c320b9feafda58d67b56ab0b80601f5bb4a1

Pith citing papers

No inbound Pith citation observations are available.