Pith. sign in

Paper Citation Record · LEDGER

Efficient Skill Discovery via Regret-Aware Optimization

As of 8 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 0 inbound Pith citation observations for arXiv:2506.21044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21044 v1

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:41:38.230312Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

74 of 74 outbound references displayed

  • verified exact2
  • verified fuzzy51
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 90c85963-0971-41b0-9aa5-55b627618389 · outbound

This paper cites write newline.

Efficient Skill Discovery via Regret-Aware Optimization write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.145979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.145979Z digest=sha256:3e7f05682d5322be76cc30f3b1e6bb0e401774fd501e9a9122db8f7e4f1cb207

Observation 1da8fa44-ed1a-4839-883e-e382ad82a665 · outbound

This paper cites M., Crump, T., and Far, B.

Efficient Skill Discovery via Regret-Aware Optimization M., Crump, T., and Far, B

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.322974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.201982Z digest=sha256:ef0cc8fc7b66847e011475974ffb86a6036135ad42831739015fcd5a4b69f005

Observation b46a063f-888c-427f-b844-28504c636e07 · outbound

This paper cites Hindsight Experience Replay.

Efficient Skill Discovery via Regret-Aware Optimization Hindsight Experience Replay

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.306850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.306850Z digest=sha256:d4011d874e449c20683495c54bcc812ff0145dc66c9d753c7af8a43fd3031c3d

Observation 88f1255f-08d8-45f4-a571-a72faee19108 · outbound

This paper cites TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations.

Efficient Skill Discovery via Regret-Aware Optimization TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.424348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.424348Z digest=sha256:acd51ed35ac0657c79bcbcfd80e71022fdb75967c875cbf371ecb99bb389a7f6

Observation 1ea52e5c-7a75-40e5-8676-473c33a22ed3 · outbound

This paper cites K., and Konidaris, G.

Efficient Skill Discovery via Regret-Aware Optimization K., and Konidaris, G

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.310533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.518360Z digest=sha256:57a77a6132edff89d0d2322ea8c9efce711a1d51fa3d38acd0d98775aa6c3e25

Observation 0e414a57-7eaf-4db7-9a17-26479fb26d17 · outbound

This paper cites Constrained ensemble exploration for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Constrained ensemble exploration for unsupervised skill discovery

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.297553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.609290Z digest=sha256:d619e31f537b5fa8394e2dc0f5919f51ceb3ed720f34795a0cad31ca09a6ef40

Observation 9a4094e2-9f68-4cd4-bdfe-ea633f5bcc5d · outbound

This paper cites X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U.

Efficient Skill Discovery via Regret-Aware Optimization X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.283764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.703614Z digest=sha256:425ad53a0e30b6568b7e5d83af7f1fd63d48fa46a5fda0e93a59cdaf03e9f34e

Observation 0f826a26-1c03-4767-ade4-ceeb660bfa10 · outbound

This paper cites Exploration by random network distillation.

Efficient Skill Discovery via Regret-Aware Optimization Exploration by random network distillation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.270252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.785283Z digest=sha256:b06a4cc9c0aa50474b090a6fc31d0ab273c0526e6b1cfae630fe3230a5fe92f3

Observation 55067e30-d11b-4a7b-9452-07042e042327 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills.

Efficient Skill Discovery via Regret-Aware Optimization Explore, discover and learn: Unsupervised discovery of state-covering skills

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.256300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.836290Z digest=sha256:d4d15cc87b95a65c9bc87c120ae4ca0bec2be90884a04381a14ba790f325bbcc

Observation 6f836df7-b49c-4975-87b2-7d70de188e82 · outbound

This paper cites Language as a cognitive tool to imagine goals in curiosity driven exploration.

Efficient Skill Discovery via Regret-Aware Optimization Language as a cognitive tool to imagine goals in curiosity driven exploration

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.241660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.917944Z digest=sha256:2b82a317e8c15c129f2d9f68b4396e915c36bf91f56635b947504bcb124bde65

Observation 0b03d062-5f28-41ad-bee5-78a4176256fd · outbound

This paper cites Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey.

Efficient Skill Discovery via Regret-Aware Optimization Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.227451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.005894Z digest=sha256:907042dfe73aa6713b9a6593901007b41fec64a100fa3a87ce755db572222e4f

Observation ba3fe812-2a9d-4457-abdc-473cae2bb186 · outbound

This paper cites Goal-conditioned imitation learning.

Efficient Skill Discovery via Regret-Aware Optimization Goal-conditioned imitation learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.212302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.090668Z digest=sha256:79f6cdd369dd82fdd74c2355c2319d87d336ebe99d927a06512467aa9e018009

Observation a5494acd-cb81-4cd8-a6e5-f246fab0f8bd · outbound

This paper cites Adversarial intrinsic motivation for reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Adversarial intrinsic motivation for reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.198143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.203128Z digest=sha256:e3e5ce9a40bffa0fa2e61c2e23c1bc024c819876611d8ed994d26e7ca5bbe957

Observation e284d72f-c053-4426-b2b3-0abe976a1ce9 · outbound

This paper cites Diversity is all you need: Learning skills without a reward function.

Efficient Skill Discovery via Regret-Aware Optimization Diversity is all you need: Learning skills without a reward function

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.184132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.247740Z digest=sha256:0a0b2a340d79a3f55dc940ff4d9743f8b1d5468d2d54984ecb80d00a4c3f2226

Observation 143187d5-3f4d-4258-9ef2-695f91139aa3 · outbound

This paper cites C-Learning: Learning to Achieve Goals via Recursive Classification.

Efficient Skill Discovery via Regret-Aware Optimization C-Learning: Learning to Achieve Goals via Recursive Classification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.332286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.332286Z digest=sha256:7b009d5eb7c89daa6cfe48d481fd0ffbaa80b3ea2843f4b1876d0d33449e5e01

Observation 20c0ea5b-129c-4409-bda9-7de5ec4a561b · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.170092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.417950Z digest=sha256:0d15623fe410d26c0e34be4391de46db27b6dee1539d480a195492b6655ce08d

Observation d9147cd4-2c24-4fa6-b982-d8eec0274326 · outbound

This paper cites Curriculum-guided hindsight experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Curriculum-guided hindsight experience replay

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.156692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.468891Z digest=sha256:8387892d450405be4d721aa4609514c6f6e4f422b0b7a574cf28a33504f16b78

Observation a5e18089-2bb6-4ebf-b2dd-aa137296a80a · outbound

This paper cites Automatic goal generation for reinforcement learning agents.

Efficient Skill Discovery via Regret-Aware Optimization Automatic goal generation for reinforcement learning agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.473619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.473619Z digest=sha256:bca5f22d309ab0679c292a1364c8e5dbc51cc451d7b4e5721bc580ba5d95b6ed

Observation 7d6cf578-631c-4d71-849a-32e0cf78ae83 · outbound

This paper cites Accuracy-based Curriculum Learning in Deep Reinforcement Learning.

Efficient Skill Discovery via Regret-Aware Optimization Accuracy-based Curriculum Learning in Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.578090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.578090Z digest=sha256:da3cde4b4dec9cee7bb0ad0902f49e9eb56ea3f618be14f7334420f51fab127f

Observation ccb63044-06ba-4f5f-8633-324719177297 · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2020.

Efficient Skill Discovery via Regret-Aware Optimization D4rl: Datasets for deep data-driven reinforcement learning, 2020

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.133429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.762057Z digest=sha256:898aa77bff2d036dbd6b5b06de4a19a5d0c4e5ff6ebf7740a788b61b352f4d97

Observation 77335ca6-e087-4292-8657-b2892f585840 · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Efficient Skill Discovery via Regret-Aware Optimization Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.963486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.963486Z digest=sha256:646853fe0f74f793ad9247460738a35ec02e3b018e0a29a3b1075e3383356bdd

Observation 53a1012f-57ab-4f05-a980-b04548088400 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.121608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.121608Z digest=sha256:f348a92bd8558ed8d0e3964b7070b6a25a2c37bc76c6c5a9ce960b6a2fbdfe85

Observation 50abf721-f062-4549-8f28-ca8c2852b4eb · outbound

This paper cites J., and Wierstra, D.

Efficient Skill Discovery via Regret-Aware Optimization J., and Wierstra, D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.110596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.261687Z digest=sha256:275b60a4377b7f0cfeaf390a786731bc8f86a5cc28ded97d036dd68201b08f6d

Observation 37668b98-4ab1-4136-be9f-db203f06ca2b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Efficient Skill Discovery via Regret-Aware Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.471301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.471301Z digest=sha256:12da9976a6e9d20f5ac6bb2bfdf7d483ebe092e6bcb740f3903c201b9fd829a5

Observation a1cbfc2f-458d-472e-a9bb-20c95be91d63 · outbound

This paper cites Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation.

Efficient Skill Discovery via Regret-Aware Optimization Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.461935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.692807Z digest=sha256:e4c4865d4c4c13678b0f4021b7f55677fb0e6c28bb00b57e68fadd57e8c5cc53

Observation 33de2148-99d3-48aa-8870-312b786ce215 · outbound

This paper cites Exploration in deep reinforcement learning: From single-agent to multiagent domain.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: From single-agent to multiagent domain

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.089016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.858079Z digest=sha256:414736f91eb59788b920205dd887cca8e1866ac407aa7a42a003407d0dd0bf73

Observation 3a774f42-b30f-4306-8538-3989effc3200 · outbound

This paper cites Open-Endedness is Essential for Artificial Superhuman Intelligence.

Efficient Skill Discovery via Regret-Aware Optimization Open-Endedness is Essential for Artificial Superhuman Intelligence

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.029860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.029860Z digest=sha256:485b6dc4eccada4093ac4e8cd11d901393bf2836b08c82485371e2ae2050c0f2

Observation ce161550-23c4-49d2-a496-0089d72ad128 · outbound

This paper cites Unsupervised curricula for visual meta-reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised curricula for visual meta-reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.074905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.228723Z digest=sha256:f55b7fc38b3782a91347f386e72749817844159cff2b264dcfc9ce7f06622d33

Observation 63d64998-cf2b-468f-9fe6-c4c89b7cb46e · outbound

This paper cites A comprehensive survey on self-interpretable neural networks.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey on self-interpretable neural networks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.398189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.398189Z digest=sha256:e9fb448cb2cf8bb810961c8ad6acb4b9f1f2beecb8675995d1a5b79eb27c6376

Observation f843fe81-8fb2-4a69-a52a-25fa443fe92a · outbound

This paper cites Replay-Guided Adversarial Environment Design.

Efficient Skill Discovery via Regret-Aware Optimization Replay-Guided Adversarial Environment Design

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:41:38.328725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.601070Z digest=sha256:b85edcead368af0c42c317b465194dd94278657f10ddcbde8f4ad126f2e91c90

Observation 1fabd924-f58b-4583-b50b-12a9223ad4f0 · outbound

This paper cites Prioritized level replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized level replay

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.061604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.669110Z digest=sha256:5b6477bbe30c4ae228202d6b5f686a68788a5d6b8a805ad717eb0b1e8b1a5a46

Observation d4f15ff1-9b76-4682-98e1-2f2825d64cbe · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.047719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.714340Z digest=sha256:3d2577ac3a7cc6328863265c4d33f3583dd5f1a7d9dd5bf63654de9ad9af1456

Observation 844b097e-fc4a-44d9-959f-3033e3aa5941 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.740268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.740268Z digest=sha256:c03b4158922cf77ecece9b55088a181d74b913cabf50e0a8591e211a58a7d7c3

Observation 3b147d4b-9840-417d-a8ce-3fcd0f63d7b5 · outbound

This paper cites K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J.

Efficient Skill Discovery via Regret-Aware Optimization K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.026333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.814734Z digest=sha256:10065d586789ef10f22f921267b73f77f07334fda95ee105069c534bad368a9d

Observation e96c04cd-b09c-4589-87e6-e8d3b65bec21 · outbound

This paper cites Unsupervised skill discovery with bottleneck option learning, 2021.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised skill discovery with bottleneck option learning, 2021

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.012568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.865430Z digest=sha256:139e158cb0bfddae3a9cf5dd4d723c0d340354dcbfb5b189e1da29507ed6c576

Observation 31c84aeb-d9aa-468b-842f-3722166e7392 · outbound

This paper cites Active world model learning with progress curiosity.

Efficient Skill Discovery via Regret-Aware Optimization Active world model learning with progress curiosity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.998539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.942087Z digest=sha256:438921a92ec9d361cd1e1520e8089230052f8ccd3200533fa135d1fdb46eeff5

Observation 661f3f1b-5f3b-4fb8-947c-aef0e73963a1 · outbound

This paper cites Exploration in deep reinforcement learning: A survey.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: A survey

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.985466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.015977Z digest=sha256:db09447a4855bbd8e8637a39997bcbcfc4212400c0425ee6aa49aab672408b25

Observation b6d116a6-a48f-4cab-8144-738129961de2 · outbound

This paper cites B., Yarats, D., Rajeswaran, A., and Abbeel, P.

Efficient Skill Discovery via Regret-Aware Optimization B., Yarats, D., Rajeswaran, A., and Abbeel, P

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.971539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.079101Z digest=sha256:82cd11c8faaa4c6642f72df1a1b8681b4c41c6e5f0e698c3187e6bfbcd6afa6e

Observation 7fb3eaea-27e2-4d7f-b4c3-a3b8e91864d3 · outbound

This paper cites and Seo, S.-W.

Efficient Skill Discovery via Regret-Aware Optimization and Seo, S.-W

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.957700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.143371Z digest=sha256:e9087c8033ceb8099ffb0834ade579c20dd7331e17963a82059b6101f683658e

Observation 95fd50be-07ad-4499-b263-8839f0ba694b · outbound

This paper cites Revisiting graph adversarial attack and defense from a data distribution perspective.

Efficient Skill Discovery via Regret-Aware Optimization Revisiting graph adversarial attack and defense from a data distribution perspective

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.944068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.200699Z digest=sha256:92e490071c1a47816ffecd9d26a05075fb20e340f0f86b2b654bd242d3d68327

Observation 5cbd798b-7508-4414-98dd-1f02158d4e1f · outbound

This paper cites Boosting the adversarial robustness of graph neural networks: An ood perspective.

Efficient Skill Discovery via Regret-Aware Optimization Boosting the adversarial robustness of graph neural networks: An ood perspective

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.929948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.273877Z digest=sha256:702d1b37ab106fd3783f4914318b06901b0e9ccfd825ba855bed58f327342ada

Observation 8972b79e-4f56-4c4d-879d-e3ff63f6d453 · outbound

This paper cites A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals.

Efficient Skill Discovery via Regret-Aware Optimization A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:37.321403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:37.321403Z digest=sha256:8374f0f906c000b926c0d3bdc51e8f221661272152c6e0b29de3ae58a3f3ab16

Observation 17c92cf3-8124-4217-8929-958f476c3f8b · outbound

This paper cites Choreographer: Learning and adapting skills in imagination.

Efficient Skill Discovery via Regret-Aware Optimization Choreographer: Learning and adapting skills in imagination

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.916681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.365884Z digest=sha256:34a1e7f63ad0b5c3f17cee302f0ab1a6257db0794d97a034f036527efcf50072

Observation 57d2ebd6-1b94-4784-b061-afed5604589b · outbound

This paper cites Planning with goal-conditioned policies.

Efficient Skill Discovery via Regret-Aware Optimization Planning with goal-conditioned policies

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.901968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.462003Z digest=sha256:977edf269d0ba84c157c2ffaaf7bcb6f031a7dc2b780bb93c07de71b48b977a8

Observation 06538529-6818-4526-94fe-164caa8d0a93 · outbound

This paper cites Wasserstein dependency measure for representation learning.

Efficient Skill Discovery via Regret-Aware Optimization Wasserstein dependency measure for representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.889465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.499235Z digest=sha256:0d7d429e73eac5d4a7af733cd798bfeab88154faa4695d1a766d9bdb0c7bf006

Observation cb1c474e-9636-4702-886b-d7f66500ee3f · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Lipschitz-constrained unsupervised skill discovery

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.876381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.582284Z digest=sha256:5ba1eaa8d7db0be5ea5199cbff7c6594ddb7aee6a020c20855bb37a54114fc6d

Observation 2318ca52-4a84-4898-b339-788fcc50afb8 · outbound

This paper cites Hiql: Offline goal-conditioned rl with latent states as actions.

Efficient Skill Discovery via Regret-Aware Optimization Hiql: Offline goal-conditioned rl with latent states as actions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.863376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.652587Z digest=sha256:21027714dd5eae327e4176cfcbe84a5f9aee452b090fbc2340099ee0a0171be5

Observation edf572e2-741e-428e-8fa2-e7cf5223703b · outbound

This paper cites Foundation policies with hilbert representations.

Efficient Skill Discovery via Regret-Aware Optimization Foundation policies with hilbert representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.849816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.697677Z digest=sha256:24d33cb6a30cc33db79e9aac64121354bfe9c471290646979247b86194483fed

Observation 1c666100-dff7-49e6-9bb4-4666000b500c · outbound

This paper cites METRA : Scalable unsupervised RL with metric-aware abstraction.

Efficient Skill Discovery via Regret-Aware Optimization METRA : Scalable unsupervised RL with metric-aware abstraction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.837059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.797162Z digest=sha256:34e82223ec41d969b62f5735c1760f05768511a7994d576757305e9c81dd375f

Observation b04ca4b0-c24e-47a8-a6e3-beb4780880fa · outbound

This paper cites Evolving curricula with regret-based environment design.

Efficient Skill Discovery via Regret-Aware Optimization Evolving curricula with regret-based environment design

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.824248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.867902Z digest=sha256:f2ee4531fa6908f6e3e4794c303fbd54dabacd3f0236dde00c77c97f110155c0

Observation 06fc72e3-eb0c-41b2-aeec-450f54d74c22 · outbound

This paper cites A., and Darrell, T.

Efficient Skill Discovery via Regret-Aware Optimization A., and Darrell, T

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.810397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.943248Z digest=sha256:789ee60817d7f291d7c9654f15486c9969189ca0d99dc3b74a7736f4c0faec5e

Observation 99fba55d-032c-4e6b-9bbb-6e0f2bb19b85 · outbound

This paper cites H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S.

Efficient Skill Discovery via Regret-Aware Optimization H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.796256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.036854Z digest=sha256:fac8b102634d106f1e44d02deeb17f8c7a713460b11cd66b7647c1cff98fc5f2

Observation 9cfa1ce2-4d30-42d0-8d0b-d85150d118b1 · outbound

This paper cites Automatic Curriculum Learning For Deep RL: A Short Survey.

Efficient Skill Discovery via Regret-Aware Optimization Automatic Curriculum Learning For Deep RL: A Short Survey

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.079354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.079354Z digest=sha256:9bb3048e8d90028c92016a4b8e4241833c18922de1639c7adbec4db42e714dc0

Observation 428819a6-0947-4e3d-90bb-107c5f699bdb · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:38.783134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.121688Z digest=sha256:886eb417c9e802e3bfedacb6ba7b75649e4042977c0301b1a73105ab5357b1f1

Observation 20125f29-6832-4238-89d4-625fbd6c8abc · outbound

This paper cites Prioritized experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized experience replay

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.770421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.131699Z digest=sha256:fb1e081ada78ef69065d162b9ce0d52667093d334cfdf55525404365d5b08757

Observation c49c44bb-ec1b-4384-93b5-5b9283d4a7b8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Proximal Policy Optimization Algorithms

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.139133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.139133Z digest=sha256:5c06486aa4309318cb9661eb08e78a3b6e4247f023337c7f4c36e6ba65762bbe

Observation 37147148-54ea-48b5-ae05-9d0211ef2774 · outbound

This paper cites Dynamics-aware unsupervised discovery of skills.

Efficient Skill Discovery via Regret-Aware Optimization Dynamics-aware unsupervised discovery of skills

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.755575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.149622Z digest=sha256:a88a7c4bfc876db0f4231f7ef5e3ffc259fac49f192af6bde016d2c6ca8bd1cb

Observation 3ece75c2-8146-4a87-aac1-db7807dc5998 · outbound

This paper cites Deterministic policy gradient algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Deterministic policy gradient algorithms

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.740517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.157683Z digest=sha256:2a928783f40cd9bddcae9e57866ee6b77b2e8eb349e87a482bbfe808d67cfb4c

Observation ae196208-c010-45c4-9dee-eea8293b97e0 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play.

Efficient Skill Discovery via Regret-Aware Optimization A general reinforcement learning algorithm that masters chess, shogi, and go through self-play

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.725250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.165986Z digest=sha256:776eba744af567b946334fcd937888a7d92c0eb3a2cf5ed508b7002dceec5826

Observation 05395d97-732b-444f-971b-fd92e2885772 · outbound

This paper cites Intrinsic motivation and automatic curricula via asymmetric self-play.

Efficient Skill Discovery via Regret-Aware Optimization Intrinsic motivation and automatic curricula via asymmetric self-play

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.710934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.170823Z digest=sha256:c337d70bf3143f76e05d7d0a548554cf9dd553f973796f9a889059273c6842b6

Observation 72c02ece-e109-43a7-a90a-62125a6fec8f · outbound

This paper cites Policy continuation with hindsight inverse dynamics.

Efficient Skill Discovery via Regret-Aware Optimization Policy continuation with hindsight inverse dynamics

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.696597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.174661Z digest=sha256:aed9aa84a4c7f9ac8a71dd84d72adef4a254cdc8fc73211d669670107c790976

Observation 9343b173-20df-48ed-9c25-1f9d65641445 · outbound

This paper cites Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections.

Efficient Skill Discovery via Regret-Aware Optimization Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.682455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.179067Z digest=sha256:852ca7fd94c3b73a63a6f87dd616164e8901953eabf26569b05469ef649d5d96

Observation fdad43a7-eb70-470d-baff-90cdd4c02170 · outbound

This paper cites Market-aware long-term job skill recommendation with explainable deep reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Market-aware long-term job skill recommendation with explainable deep reinforcement learning

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.667895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.183587Z digest=sha256:9d4d5c12678a6af1776d1d7e767add0284d437d2d48df10ca2a4066fce3f76ef

Observation 8abb6957-0f84-4e6a-b131-b2d2787a6ba8 · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.188231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.188231Z digest=sha256:f0ebf1d3f430fcf13cc856741afa026e69c572e4d10fb43e29db284b29757324

Observation 2c1ae144-c4ff-4eb7-bd6d-223d2bb13244 · outbound

This paper cites S., McAllester, D., Singh, S., and Mansour, Y.

Efficient Skill Discovery via Regret-Aware Optimization S., McAllester, D., Singh, S., and Mansour, Y

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.191955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.191955Z digest=sha256:c3bf7bbe7214510ebdd8886e708babc25bb101d02eb1c3f1a633f78a60ccad72

Observation 0c85fde9-bacb-4a1b-8a53-ba00562a2777 · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Efficient Skill Discovery via Regret-Aware Optimization dm\_control: Software and tasks for continuous control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.637023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.196846Z digest=sha256:72c2bbd9247b02ca2b81a477fa4b2363b071406f2d76110607ed0f220cd4a23f

Observation 7fe00edc-7e18-485c-8903-9a3394f52bc8 · outbound

This paper cites M., Mathieu, M., Dudzik, A., Chung, J., Choi, D.

Efficient Skill Discovery via Regret-Aware Optimization M., Mathieu, M., Dudzik, A., Chung, J., Choi, D

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.623333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.201024Z digest=sha256:9c88854499b6ea485f1d3db0255b1edb521f601d1aae6c39beff97b20331d295

Observation d7b4d0c5-2d5f-4bdc-84e9-52b926ab9d66 · outbound

This paper cites Optimal goal-reaching reinforcement learning via quasimetric learning.

Efficient Skill Discovery via Regret-Aware Optimization Optimal goal-reaching reinforcement learning via quasimetric learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.608881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.204893Z digest=sha256:58114d85e992e7606704b5b498acdc6a2ad1b504cf11e1f6c1615c77e120b762

Observation 366bb4e3-00e8-4f0d-9c7f-15d14fc0b0d7 · outbound

This paper cites Ski LD : Unsupervised skill discovery guided by factor interactions.

Efficient Skill Discovery via Regret-Aware Optimization Ski LD : Unsupervised skill discovery guided by factor interactions

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.594975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.209112Z digest=sha256:a0fe68787f3b943fbf0da694e007dc291fd8aebdc7d997111dbd883fd5f3a2ee

Observation 14c24df4-3b5b-43a6-8166-d494a170eb41 · outbound

This paper cites A comprehensive survey of forgetting in deep learning beyond continual learning.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey of forgetting in deep learning beyond continual learning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.581685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.213181Z digest=sha256:00837d6734d6a318f1e0c943f298187b33e0dc9fc4f1a89ae87decfa5507e835

Observation 2d352df8-1244-410b-94cb-5d6481ba4334 · outbound

This paper cites Neural Program Synthesis By Self-Learning.

Efficient Skill Discovery via Regret-Aware Optimization Neural Program Synthesis By Self-Learning

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.271036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.217400Z digest=sha256:6b61bc36ef350f8ed61f83a95419e0eacfab369ba315c7c6d0ba4e1393c0a057

Observation 499e79d3-9c9e-4f05-ba46-034c0a95d1b5 · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Behavior contrastive learning for unsupervised skill discovery

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.568005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.222361Z digest=sha256:a86ea4d7464b7471f6ae3c8b8f2a71147d92260904195752dcdb1fcb17fc5bb7

Observation 9db995b6-e3eb-4e5b-9d9b-30ce1be6506b · outbound

This paper cites Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.553086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.226237Z digest=sha256:97ddfede6300b1f33a329653f241e0563aa89024b58297f9af4589b87e75fc87

Observation 6c028a05-0532-41b4-8993-8abd49123dec · outbound

This paper cites Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach.

Efficient Skill Discovery via Regret-Aware Optimization Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.538713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.230312Z digest=sha256:e3ee149198a2ff0b4270367a208373d1a762dafd1f7b4ed8cdc8686e9b63b94e

Pith citing papers

No inbound Pith citation observations are available.