Pith. sign in

Paper Citation Record · LEDGER

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation

As of 24 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2411.16532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16532 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:04:33.476637Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy28
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3d6f0111-0dc4-4dad-ac00-972585748772 · outbound

This paper cites Cells 10(4), 735 (2021).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Cells 10(4), 735 (2021)

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:34.000303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.302448Z digest=sha256:2d4f068171930aceabaee55785fb3f0bff9e5fc249fb88cc90903eed5e9cb365

Observation 5ab6beb6-2e3c-4c75-aa1c-f29ddc6b8f8b · outbound

This paper cites an unresolved cited work.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:04:33.987797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.306664Z digest=sha256:4676d68b0b75a21f325c7b65708437f3aee5e93aea66e3ac862db21a6c0a63d3

Observation 508ab118-4129-42c4-9261-d7d65c13c3b5 · outbound

This paper cites Science 245(4918), 605–615 (1989) 30.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Science 245(4918), 605–615 (1989) 30

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.975964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.310553Z digest=sha256:6851f33a31f5305132c9acd0c64b05bfccf8a583835ca5e1bbc78bfbf90651e5

Observation 1b08ebbc-5438-45d4-95fd-863080473ee9 · outbound

This paper cites Neural networks 113, 54–71 (2019).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Neural networks 113, 54–71 (2019)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.964520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.314578Z digest=sha256:cd72e63958b99c0e0c258c6f84899b0b4ef81254d48da316d101cfb3783d6e2a

Observation 7cee9194-7588-4155-8b8d-bcf239d5864d · outbound

This paper cites Journal of Artificial Intelligence Research 61, 523–562 (2018).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Journal of Artificial Intelligence Research 61, 523–562 (2018)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.953167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.318584Z digest=sha256:cc9cb5f2ed7b61f661a4e84b41bd8e94328f26b67044b0637fd9777e4f1d76b9

Observation 8bc7dc0f-5713-488e-97aa-023518d0b04b · outbound

This paper cites To Compress or Not to Compress- Self-Supervised Learning and Information Theory: A Review.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation To Compress or Not to Compress- Self-Supervised Learning and Information Theory: A Review

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.323479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.323479Z digest=sha256:edd0df54521ec531b45c319176baec41c23335fbde5dd76d7963fb28de6e0fb2

Observation 59fa3a3f-e423-4c6b-916e-96f4b5a60a01 · outbound

This paper cites Curiosity-driven Exploration by Self-supervised Prediction.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Curiosity-driven Exploration by Self-supervised Prediction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.328097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.328097Z digest=sha256:997e97f550ec7143b1e8c58682e0d8e30a00781c8ec1272512e9d1bb26303399

Observation bfe4ed04-c549-4203-957d-74e71ba5c79c · outbound

This paper cites Advances in Neural Information Processing Systems 34, 20516–20530 (2021).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in Neural Information Processing Systems 34, 20516–20530 (2021)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.942572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.331989Z digest=sha256:4689f7c25191fc90dca4ebee95a6039e43146d1ebf35c56eb3350a9932c3f25a

Observation 13fa37b2-2ff8-49f2-a231-366d5667748f · outbound

This paper cites Advances in neural information processing systems 12 (1999).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in neural information processing systems 12 (1999)

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.335646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.335646Z digest=sha256:4d64b1fa2fe67dc9737595bf84b20eec5d6bcaa2bde4275c1c8d62c6e11a41d5

Observation c1019052-cfe8-4249-957e-95ac70dc885a · outbound

This paper cites Advances in Neural Information Processing Systems 33, 11734–11743 (2020).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in Neural Information Processing Systems 33, 11734–11743 (2020)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.925079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.338866Z digest=sha256:f6de283b2de7d53e7becb4895068760002d67245f8146c68bf51b3dc92ba1694

Observation e668401b-a5e1-4d5d-b45c-9c77a738700e · outbound

This paper cites Proceedings of the national academy of sciences 114(13), 3521–3526 (2017).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Proceedings of the national academy of sciences 114(13), 3521–3526 (2017)

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.342505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.342505Z digest=sha256:9e7b13ae94eaa01674b484251392a8189cfbb446264f7f6f17ec3c5ff3aa6cb7

Observation f9854f95-356f-43d3-85b2-5734f2bd88bf · outbound

This paper cites Progressive Neural Networks.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Progressive Neural Networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.346040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.346040Z digest=sha256:abd9f1dac40fc5759aa95f7cd906ccbc5238aa85c824169480573b0abdabaca7

Observation e2910c15-50c8-489d-99e8-d097661aa008 · outbound

This paper cites In: International Conference on Machine Learning, pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: International Conference on Machine Learning, pp

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.907919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.349859Z digest=sha256:e290c9e01d55fa5b493795e9fc567e601d35f9c04bd1125c47160025f894836e

Observation b04eb15e-0154-4f2b-9c76-8394f0ab9f72 · outbound

This paper cites Task Agnostic Continual Learning Using Online Variational Bayes.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Task Agnostic Continual Learning Using Online Variational Bayes

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.353425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.353425Z digest=sha256:d1a69699e790afa9688021c7725294d9675f043dcdf9317b8855c3ea0f79fac6

Observation 4dbb8b79-7737-40ae-90da-92c40a2cd1d4 · outbound

This paper cites In: 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 31 pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 31 pp

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.897375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.357158Z digest=sha256:a8e94624d6a549110dc20f6ea7cd644ccdcd0bd23992635d9be573ece6a3aa94

Observation d4a4c9f3-7bd1-4092-8ec8-7a5b96d9b95b · outbound

This paper cites The Arcade Learning Environment: An Evaluation Platform for General Agents.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation The Arcade Learning Environment: An Evaluation Platform for General Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.360905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.360905Z digest=sha256:5ef70dc6e96baed9543c71c662e2cc171b6b702076ec67c0e2436db8aefac1f8

Observation 0961907f-0f71-4b7b-ad9b-a7d81d5bc46c · outbound

This paper cites In: ICLR (2016).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: ICLR (2016)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.886216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.364692Z digest=sha256:080b9c5a276d40b23cdee708bcb022e0af2de39b0fbf325a3f030d9faa3601ca

Observation c5d5095a-d4ef-413f-b0f0-087cb09207d1 · outbound

This paper cites In: The 22nd International Conference on Artificial Intelligence and Statistics, pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: The 22nd International Conference on Artificial Intelligence and Statistics, pp

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.875748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.368406Z digest=sha256:2303f561e2199acfbd7e13521150f26357c5e4f49bb1ba0d25da36e8388994a3

Observation df9801a2-ea65-4026-ad22-c81c6ac4f70b · outbound

This paper cites Advances in Neural Information Processing Systems 34, 6920–6933 (2021).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in Neural Information Processing Systems 34, 6920–6933 (2021)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.863519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.371848Z digest=sha256:844cfb1ff8f293520113bef5a9883352d4d6840ff43035b89d0347cefa568a46

Observation 3f39f943-d78a-4820-8275-10afc997d5ca · outbound

This paper cites Real-time Policy Distillation in Deep Reinforcement Learning.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Real-time Policy Distillation in Deep Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.375549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.375549Z digest=sha256:31bc0158d408a55ca082fad88ae93edda955db909f023ca041de9fb586fd708d

Observation 0a932111-dc06-400f-801e-52eee6581911 · outbound

This paper cites In: Proceedings of the 16th Conference of the European Chapter of the Association for Computational Linguistics: Main Volume, pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: Proceedings of the 16th Conference of the European Chapter of the Association for Computational Linguistics: Main Volume, pp

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.851039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.379369Z digest=sha256:00eb0e72080397520042f558c3bd8a7974083d8ec1e4721699c8ae4194cf4615

Observation 200a8345-0f36-4f0e-a2ea-ee66be6519cc · outbound

This paper cites In: 2017 International Joint Conference on Neural Networks (IJCNN), pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: 2017 International Joint Conference on Neural Networks (IJCNN), pp

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.838922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.382803Z digest=sha256:015c1dddef8adc0ad78550b0cf9534266fca1f266f921f9977dc6ad5866c7dad

Observation 18edfbd2-871e-425a-9738-021deae3083d · outbound

This paper cites IEEE Transactions on Cognitive and Developmental Systems (2023).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation IEEE Transactions on Cognitive and Developmental Systems (2023)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.827785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.386277Z digest=sha256:9ba03b495d8359eb492d0a141b01fae90571707eedf6d069a6c0fc17cdbf089f

Observation 6f89f528-1589-44b8-840e-43b2b44fee02 · outbound

This paper cites In: International Conference on Learning Representations (2018).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: International Conference on Learning Representations (2018)

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.816581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.389631Z digest=sha256:9b8ae30bcee86562b5f68c15b79664af5f381e3d65a1da21e96968ade69154ce

Observation 3a55fb09-7745-44f9-afc3-2754b00c7c39 · outbound

This paper cites Nature communications 11(1), 4069 (2020).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Nature communications 11(1), 4069 (2020)

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.804356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.392888Z digest=sha256:952abfe6bd1bc221ab1f9532b0b5b90555b8028673de5663356443530bb4eb98

Observation 72f21cca-6515-4f37-88c2-7312c816baee · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence (2024).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation IEEE Transactions on Pattern Analysis and Machine Intelligence (2024)

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.396185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.396185Z digest=sha256:6aed9f9cfdc2c647ce146a0dab2ad162d3afc0be862924e19276f422c810b5b4

Observation 84193407-894a-4158-a799-733714799c99 · outbound

This paper cites Advances in neural information processing 32 systems 32 (2019).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in neural information processing 32 systems 32 (2019)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.787414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.399706Z digest=sha256:6d3682ee84d078a7dfdc74a4abb9c99b8f9fc240a02719228fc3f3da0dba8a83

Observation f4fae2a1-5c9a-434d-bac5-5f9cb3089932 · outbound

This paper cites Representational Continuity for Unsupervised Continual Learning.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Representational Continuity for Unsupervised Continual Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.403044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.403044Z digest=sha256:2f3311ddcf6b35b6a8bd9ab2552f30e2fb377aa96bc5b994ea009eabbaebc010

Observation 82020a5a-f4a8-4ade-8d5b-ebe7363f646c · outbound

This paper cites Journal of Neuroscience 35(3), 1319–1334 (2015).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Journal of Neuroscience 35(3), 1319–1334 (2015)

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.775722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.406484Z digest=sha256:f703c3c31f85e53f63f4b17d708b2ade8650ebd6fa902c8e3867ee2bfe51fe6d

Observation 466d7de1-1ff0-461e-8bea-d9da57b2bba7 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Distilling the Knowledge in a Neural Network

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.409953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.409953Z digest=sha256:c144f0b3d350855fceab02b1335980dd4c52ab4b023fde4c3cc7b7e47da94160

Observation e3b32040-905c-48e9-9248-3a85e909a8f8 · outbound

This paper cites Advances in neural information processing systems 17 (2004).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in neural information processing systems 17 (2004)

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.764965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.413941Z digest=sha256:dd4f9147fdaaa4f7f0796b273425f729a95e7c6f70830724cec8a767de1669fb

Observation a73f815f-afdf-450b-8c0a-4496db8d9564 · outbound

This paper cites Neuron 36(2), 285–298 (2002).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Neuron 36(2), 285–298 (2002)

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.754102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.417530Z digest=sha256:7efedef44a9b218e49001b5ccc239209386ba4555275064e2d44eb3ffe240fc3

Observation f1adc505-7f3d-43bb-be03-3b96eb5aca1b · outbound

This paper cites In: 2017 Joint IEEE Inter- national Conference on Development and Learning and Epigenetic Robotics (ICDL-EpiRob), pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: 2017 Joint IEEE Inter- national Conference on Development and Learning and Epigenetic Robotics (ICDL-EpiRob), pp

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.742826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.420796Z digest=sha256:155a678c5787ee6517fa7ef2da2f5a740cfafd143a44e35d51c39910eae1b395

Observation 24e8eeb7-bb33-4bb6-98ce-8d77ec2cb8aa · outbound

This paper cites In: Proc.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: Proc

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.730510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.424393Z digest=sha256:549de9db2b09ba73c8f3bdf807e0b3a0efed45e604c64dab82350bcc5efb9ff3

Observation f2d877f1-c5a5-41d2-9ef3-d808eded0a87 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Playing Atari with Deep Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.428019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.428019Z digest=sha256:da57f0f8faf92cbf478335443499f3ae58d660ef38486652d1bc2911f827f4aa

Observation 903530e3-5c2a-4613-bf72-072787bbcfcb · outbound

This paper cites Deep Reinforcement Learning: An Overview.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Deep Reinforcement Learning: An Overview

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.432291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.432291Z digest=sha256:4c517c8aadeaf634fe31759c9069fd000bb605652c2e91e1dfa4fa469e8f1ff2

Observation a4727dc8-b876-48ca-ad64-1e0c6ff400f8 · outbound

This paper cites In: International Conference on Machine Learning, pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: International Conference on Machine Learning, pp

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.718636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.436012Z digest=sha256:70d2486d27135b6dd6b95d9b31ee3abb30b99bbd11ed81b71a4031e018bc5e03

Observation 761d1af0-06e1-459c-8e8c-a7cd51bb8761 · outbound

This paper cites Advances in neural information processing systems 30 (2017).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in neural information processing systems 30 (2017)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.707913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.439530Z digest=sha256:dfe9150fc014e243fabe1b55137c1662bb97b9840621d667e53f170de1c2a329

Observation 2f092a20-d89f-4343-9afd-1b09ddd698ce · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.442693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.442693Z digest=sha256:b4ba6dadad9bd45c23d1ba2f4e8db74b9d3849cda3116b69a8a3ddd569402783

Observation b4e07829-0e75-460c-b752-d65fbe8009a0 · outbound

This paper cites A Comprehensive Survey of Continual Learning: Theory, Method and Application.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation A Comprehensive Survey of Continual Learning: Theory, Method and Application

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.446520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.446520Z digest=sha256:265af8d96cc5ca308680523f0910dcaa00e4d2d694444f62fa226ee52a5a088f

Observation 963beb16-2806-4317-ae76-541026cf762e · outbound

This paper cites The Royal Society (2017).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation The Royal Society (2017)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.696586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.450381Z digest=sha256:a277f1e40f8e977ce949915aa8ef50858b120d252a6f73dfff0ed08248ae44f8

Observation 6b8dbcde-d17d-481b-880b-40507494b369 · outbound

This paper cites Three scenarios for continual learning.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Three scenarios for continual learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.453841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.453841Z digest=sha256:cfff9db6e0415dccedb415d0af59374d589fbf9d407845e8135cf71cf51b9908

Observation 14362b97-6061-4366-a6b7-176c781f394f · outbound

This paper cites In: 2020 IEEE Conference on Games (CoG), pp.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation In: 2020 IEEE Conference on Games (CoG), pp

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.685442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.457913Z digest=sha256:c75e70803f2e19a61acf9fbf7382722114aedbf8af18bc13a69f46c01a1b25f5

Observation 32a9bff2-f1d3-4806-89fd-485e27cad26a · outbound

This paper cites Advances in neural information processing systems 32 (2019).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Advances in neural information processing systems 32 (2019)

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.461674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.461674Z digest=sha256:135ba46a51339a5b6fa545c282d1dc17e54ed4fe2d8e5a1d05d351402a42dbf2

Observation 357f10ae-bf2a-469c-a1a8-98053f6a1c91 · outbound

This paper cites GitHub (2018).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation GitHub (2018)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.667839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.465196Z digest=sha256:4465c103d6b09bcc5f1295b9aff6698643ea2e3aaf48fff5c81f6ddbf1589143

Observation 13d9c609-195f-4662-bbdf-d02a6844f5fb · outbound

This paper cites Journal of Machine Learning Research 22(268), 1–8 (2021).

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Journal of Machine Learning Research 22(268), 1–8 (2021)

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T13:04:33.469035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:04:33.469035Z digest=sha256:cacdbd218fc4d36ac05b0e039c3279222cd5bf74975f8a6422251a5a1931dceb

Observation b0eab509-9570-4048-9417-1017022ca7ef · outbound

This paper cites Avalanche RL: a Continual Reinforcement Learning Library.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation Avalanche RL: a Continual Reinforcement Learning Library

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:04:33.511484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.472852Z digest=sha256:12eb0be7358d08062f20daeb3355b4f2fb35a839bdd13cb5f577f59662af0433

Observation ef5160b8-72b1-44a0-9221-c0237abc6824 · outbound

This paper cites IEEE transactions on pattern analysis and machine intelligence 44(10), 6715–6728 (2021) 34.

Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation IEEE transactions on pattern analysis and machine intelligence 44(10), 6715–6728 (2021) 34

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:04:33.650548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T13:04:33.476637Z digest=sha256:157b29ed88f03f8ddaf977cbb9670b83c4cc52712f5aba5b0a23b04084fe0bd2

Pith citing papers

No inbound Pith citation observations are available.