Pith. sign in

Paper Citation Record · LEDGER

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2506.05716.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05716 v2

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:50.823234Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact3
  • verified fuzzy13
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 33b4becb-6c3c-484f-b044-5671c5f06bac · outbound

This paper cites Averaged-dqn: Variance reduction and stabilization for deep reinforcement learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Averaged-dqn: Variance reduction and stabilization for deep reinforcement learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:56.131377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:47.898294Z digest=sha256:004fb0af5cd8cabf6e5f24df2b551c2050cd360392d619ae68d57761cce097cc

Observation bcf78f58-43f0-4ac9-b332-8160d5e4c52d · outbound

This paper cites Ensemble Reinforcement Learning in Continuous Spaces -- A Hierarchical Multi-Step Approach for Policy Training.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Ensemble Reinforcement Learning in Continuous Spaces -- A Hierarchical Multi-Step Approach for Policy Training

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:18:51.769542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:48.008298Z digest=sha256:3e2511a489f22488777b9ea9412157919e5dc1bc695e2a4ff790d47c11bd5006

Observation 14e3c160-e51c-4969-9fc3-440e199c68fb · outbound

This paper cites Mushroomrl: Simplifying reinforcement learning research.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Mushroomrl: Simplifying reinforcement learning research

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:55.709013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:48.186274Z digest=sha256:92f4599cc322526c2cb2187db7a342c0f43b15e6b9df25b6723590b1dd5fa055

Observation 2b222622-40b9-4bb9-8135-53c5aa2cfb3b · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Addressing function approximation error in actor-critic methods

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:55.310590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:48.367202Z digest=sha256:b2ef018b5227e51a59b32dd2d6a9a59871af5faf2542bce35e56b87706b43ae9

Observation b62061ba-c9dd-42b8-aaed-bf34cb00963d · outbound

This paper cites Understanding Multi-Step Deep Reinforcement Learning: A Systematic Study of the DQN Target.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Understanding Multi-Step Deep Reinforcement Learning: A Systematic Study of the DQN Target

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:48.523669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:48.523669Z digest=sha256:447d29b0ab52654f890999926e539203ebcc09a142ac3fa8950d9dcbde15093f

Observation 174abb56-8cf2-45ff-ad77-9f6507898be5 · outbound

This paper cites Wd3: Taming the estimation bias in deep reinforcement learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Wd3: Taming the estimation bias in deep reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:54.888632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:48.656883Z digest=sha256:ec5a2f2103d20353e948bcc3edf1bb594f665dd8b254abad5d346a685ab3feec

Observation f2c2aaa1-ae24-4f13-9962-2d9aa4d8231a · outbound

This paper cites Rainbow: Combining improvements in deep reinforcement learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Rainbow: Combining improvements in deep reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:54.553122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:48.816458Z digest=sha256:3f1f462de829b52d567a385e5a7223a0a54198ecf691f4a9a4340aff4f8575ba

Observation bfd774fa-c87d-4041-852c-1d176ad83b9c · outbound

This paper cites Bring color to deep q-networks: Limitations and improvements of dqn leading to rainbow dqn.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Bring color to deep q-networks: Limitations and improvements of dqn leading to rainbow dqn

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:54.200244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:48.972152Z digest=sha256:15f8a807c17b403d84ad4c4a9e74607dcb32bb1eae4201e1b8063a5e78e1d552

Observation 46c982f7-5272-484e-986c-c405eddb4eb8 · outbound

This paper cites Elastic step ddpg: Multi-step reinforcement learning for improved sample efficiency.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Elastic step ddpg: Multi-step reinforcement learning for improved sample efficiency

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:53.967998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:49.072592Z digest=sha256:ed238de23f2b90d8bb937e59e3a4ec68eae1219652ca3b8f4256271693e3cef8

Observation 2b808023-ad6d-4c8d-86c8-875a23d6a104 · outbound

This paper cites Elastic step dqn: A novel multi-step algorithm to alleviate overestimation in deep q-networks.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Elastic step dqn: A novel multi-step algorithm to alleviate overestimation in deep q-networks

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:53.652559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:49.251341Z digest=sha256:2a6ad9542b04a5788e6c8ba1de9d5455360b67cbd8077b4976018ded32d2ec4e

Observation 9086a3f5-1f5e-405f-a0a9-1042649faf05 · outbound

This paper cites Maxmin Q-learning: Controlling the Estimation Bias of Q-learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Maxmin Q-learning: Controlling the Estimation Bias of Q-learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:49.430393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:49.430393Z digest=sha256:fb7930a0eb0cf1940a863aec1bde1c988cf420ca5a36f7ec7e5324ec77a7637e

Observation 78c45b44-0ade-4be2-9f43-d2709e2006af · outbound

This paper cites Human-level control through deep reinforcement learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Human-level control through deep reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:53.280170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:49.568478Z digest=sha256:01bc39591abb6d8507186c93dab9562f79cbdd855a0768ef3be7ced4aa4e8596

Observation 77e2d3d5-0248-4d0c-98b1-276b66674e6a · outbound

This paper cites Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:49.734756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:49.734756Z digest=sha256:5f2c5decc52883126069befc79990cea1f2de06b776bd869eb5b22fec021a584

Observation f35cc21c-bff7-485b-ae2e-f58f3538ba08 · outbound

This paper cites Revisiting Rainbow: Promoting more Insightful and Inclusive Deep Reinforcement Learning Research.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Revisiting Rainbow: Promoting more Insightful and Inclusive Deep Reinforcement Learning Research

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:18:51.461428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:49.846485Z digest=sha256:440a07c6ccec44889e85d469070e33056f12ca7e04df7bc8a0d402b606176505

Observation e5634b6a-8ec5-4164-aa8e-5ebcbc7f9d8e · outbound

This paper cites Issues in using function approximation for reinforcement learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Issues in using function approximation for reinforcement learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:52.984969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:49.976666Z digest=sha256:3c4cfd4bc307864081c6e1f843e582662bb24b30a762ff3b54c0d255a5476a87

Observation ba0a0d1d-eb7c-4e1d-bf74-8d37a9cbc8e3 · outbound

This paper cites Deep Reinforcement Learning and the Deadly Triad.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Deep Reinforcement Learning and the Deadly Triad

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:50.077455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:50.077455Z digest=sha256:f4d62a23ae64b6460eb86111893763ac99b9b1600dfb4cfba5f1b8833bbaf3b6

Observation 299c696d-4d85-414d-a6f9-e422f2d815ce · outbound

This paper cites Deep reinforcement learning with double q-learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Deep reinforcement learning with double q-learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:52.691097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:50.210042Z digest=sha256:fb636aa9b46ad62b11915e1fd2ac21b1c785f3e4d674f15a61481b617bdf2975

Observation 137e1cfe-eb05-4c37-9e30-8732445e98ce · outbound

This paper cites Controlling underestimation bias in reinforcement learning via quasi-median operation.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Controlling underestimation bias in reinforcement learning via quasi-median operation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:52.398852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:50.365061Z digest=sha256:d8cfc24c805e6b7516584a5fc39d3d9b8293de535d6cb6e33c4673fcfebeb26d

Observation eb2d4639-c01a-4524-8ed4-ddb292c4ae87 · outbound

This paper cites MinAtar: An Atari-Inspired Testbed for Thorough and Reproducible Reinforcement Learning Experiments.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning MinAtar: An Atari-Inspired Testbed for Thorough and Reproducible Reinforcement Learning Experiments

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:50.501864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:18:50.501864Z digest=sha256:844cae60b0661b908cdee2bf031def6c98d477ebacb595bc37a7a88e45a49c9c

Observation a6c0836e-6e67-4db9-82e9-9a314c87e66f · outbound

This paper cites A novel multi-step q-learning method to improve data efficiency for deep reinforcement learning.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning A novel multi-step q-learning method to improve data efficiency for deep reinforcement learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:18:52.106309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:50.627750Z digest=sha256:3de4a07d21e48cbca799df8623f644c12a4bbcc8ee9c77cb3cc96e379c51c1ff

Observation d80f9e1b-7935-4a6e-b508-740b9734ffdb · outbound

This paper cites beta-dqn: Improving deep q-learning by evolving the behavior.

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning beta-dqn: Improving deep q-learning by evolving the behavior

Reference 21

Resolution
verified exact
raw_fallback, observed 2026-08-07T10:18:51.148670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T10:18:50.823234Z digest=sha256:1a7175d8867f0ddc93d17b4145a18668a95d9f426fa0c822d4d9807bc036c74a

Pith citing papers

No inbound Pith citation observations are available.