Pith. sign in

Paper Citation Record · LEDGER

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2506.13672.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13672 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:32:54.400821Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact9
  • verified fuzzy21
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 088c0f74-fe5f-4416-898a-0d8fa0a6237d · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.191659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.034330Z digest=sha256:6767c548c0a93c20e4652c37f0c4f51eb68c4aecc3cbdf98902d63c3c3035e68

Observation d8740eef-1ef0-4a05-a088-89479bf979f8 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.182111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.070341Z digest=sha256:2f80c184e57fbfa16e0fd33930b1edfbbe354b722047d2f87b78bbbd087aec85

Observation ba96e2ff-57cd-4c4e-be56-0657917e6209 · outbound

This paper cites OpenAI Gym.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.109869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.109869Z digest=sha256:5016403e40b8803f8c7d769c038ec99a71d8615c0e6b75d6411af6aaee08ab00

Observation c78b0f73-f7eb-439d-9fa3-82d439d71346 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.172292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.173653Z digest=sha256:257eeb218756236880d02790e440e24d13ca11e677dd392f5640803336d34b49

Observation 94e30e18-13d6-4818-b1e4-163f389d09a6 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.161746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.232671Z digest=sha256:98bac32e3faba302c94346c4f0a0a7f30f6ade594a6a85eee6141fb29c936b93

Observation a084e66b-bf25-4ca5-b7ef-e03e92b9180e · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.151864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.284609Z digest=sha256:a8eacc2db6e0881d736bdd2b45c0ff19f4f1268932de3cacb8a23edd72a0ba10

Observation ac4175c7-aa18-49a8-8509-72ac5a37170c · outbound

This paper cites Stabilizing Off-Policy Deep Reinforcement Learning from Pixels.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Stabilizing Off-Policy Deep Reinforcement Learning from Pixels

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.368740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.368740Z digest=sha256:1d639522df71935a4b371458d5e186a8f63130cbdcb8fcb84d112b5db6b33f61

Observation b58b6687-ec6e-4d18-b8aa-6fc75da5b71c · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.412668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.412668Z digest=sha256:528aa080777c3f9d7598a5f2c3f3c0f44c12093432d311944ea8fa853a0f414d

Observation 47f52821-2c7f-4b10-a141-14f7bd75a73f · outbound

This paper cites Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.773730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.489700Z digest=sha256:8dae30aa146b4317dc8297b845179715d5bd9cacbdfefa1608d6d60d8fae5050

Observation 134041a6-9e97-48bb-8d6f-ba2341c7ff5f · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.925914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.552848Z digest=sha256:6456f478a29ef0fa5bc8d0a529245e316d460f03f2f03a191d99ebebd1440aa8

Observation a8db4646-1acd-404e-87ea-702597f337cb · outbound

This paper cites G., and Courville, A.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning G., and Courville, A

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.592830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.592830Z digest=sha256:c49e55341f21b09b2eaf98d9c10db43f6bb29b9e65f5aa6865d5d1606d9b08c7

Observation 74eb255d-4dae-4ba4-9b9c-1872bf0fcfcd · outbound

This paper cites W., Subramanian, J., and Ghassemi, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., Subramanian, J., and Ghassemi, M

Reference 12

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:32:55.758573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.628887Z digest=sha256:64b5d2e5530d34d6ac9ad81166f2241acae179b31fd5fea9de0ba6ef606c9d2b

Observation 67b3ad1d-71b8-4f47-973f-7ce034d107c7 · outbound

This paper cites Revisiting fundamentals of experience replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting fundamentals of experience replay

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.703013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.703013Z digest=sha256:fb100d3369c4da556dd5a22b4d9b12c43210895f69e08e78e92a43f3d8044250

Observation b23c158e-46a9-4aa9-b942-a735050d9861 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Addressing function approximation error in actor-critic methods

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.764866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.764866Z digest=sha256:f09e3ef9280c21a2e6d8e1e7d3a2f2f8f93b26830bdd1dffd1f8b42a1ed87328

Observation c08bb5d9-d454-446f-b43f-21835d58dad6 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.832637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.832637Z digest=sha256:f17a0536c62d2b6a2db339b8c8bd05b1e0d23c0bfc954da1bfeb22a0f495bb2b

Observation 61054b92-983b-4b0b-831e-0dc697467c82 · outbound

This paper cites Sunk-cost fallacy and cognitive ability in individual decision-making.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sunk-cost fallacy and cognitive ability in individual decision-making

Reference 16

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.706719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.905906Z digest=sha256:7c357447ac431758fcef5df71344256dd96d470fdbe237d0117e44df59167511

Observation 234a0dae-0fd4-42b5-a57c-451e32a7c932 · outbound

This paper cites R., Millman, K.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning R., Millman, K

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.116543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.926428Z digest=sha256:1adcde93c543c4bc8b432c1dc724a10c16f034040221ea19530db10d012987bf

Observation cf7f6777-c012-435d-ab6e-bb8e2237927d · outbound

This paper cites Double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Double q-learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.107696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.025987Z digest=sha256:6d9c52d7b2c3f4d5697acaca6324e5fa0aa04294a62cecf54de8a4412aea06a8

Observation 962346cb-77aa-4e31-a6e4-c439b653c794 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.089069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.089069Z digest=sha256:0f7fa87d5f7746d19c233ce3ffb7203181ec6f0433728189b7e0f205dc8746f8

Observation 30c170d8-8f14-4a6c-8741-82a3045d317f · outbound

This paper cites Planning Goals for Exploration.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Planning Goals for Exploration

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.133496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.133496Z digest=sha256:f96ca4e2fc2167173219ec9242d65479af75f805270f864bb7bb77c21ce3bd3c

Observation 0b6767b6-f2ed-4dfd-8933-e080872a10b3 · outbound

This paper cites Enhanced Experience Replay Generation for Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Enhanced Experience Replay Generation for Efficient Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.201826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.201826Z digest=sha256:4ee51f8ea90147fb8f9c9f8e390a99a9c33f7d80ce51af7a21626626b831de95

Observation a7febbd2-a0e6-47f6-8055-a046938b0574 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.098144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.236555Z digest=sha256:eb55dadda535da617480e5e13f856673adcc12cb3bbc90df5ceff2360cb78619

Observation a28efe13-186d-491b-926d-7db7514d8bb1 · outbound

This paper cites An investigation of generative replay in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning An investigation of generative replay in deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.089516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.264341Z digest=sha256:94ebb0a3d4450a415be0794e40cedc6b87809a9066e71dc5b5da2586588945ec

Observation f8429033-e01f-4244-9be0-55bba015ae5e · outbound

This paper cites Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:32:55.635890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.310272Z digest=sha256:334f036ddf8e569a1e381c0b88988901c2d9759345cf64b453087fbb100de04b

Observation f931b24d-da63-42ea-8c8b-e07459ba2322 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.372728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.372728Z digest=sha256:67c7a1b33b59454bfe7297d4ca1445e503ae8e6a8a11f172c76afff02d9c426b

Observation 7f149a22-56a5-47e0-b243-32fd713826ff · outbound

This paper cites CURL : Contrastive unsupervised representations for reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning CURL : Contrastive unsupervised representations for reinforcement learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.080060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.422423Z digest=sha256:5a70510bb80dd880faa13b27f60c4e8d6759eeee5ca2534d2357d8ceb11f6219

Observation 4eba7925-a8d6-498a-bb0c-2245d0a24203 · outbound

This paper cites HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.484896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.484896Z digest=sha256:e9f51cd7b6675954c3427e894a8c457517e75c1783afea128f0d5d040d77f40c

Observation 9027ad0e-f8f2-4657-a7be-1ecfbd57f57a · outbound

This paper cites Continuous control with deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Continuous control with deep reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.581844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.581844Z digest=sha256:c662d6091be378974a003ca54235e0920b28a948cf25941d7a92b431608f58e8

Observation 8ca1f07c-c3b9-4933-8c83-ec8274f9d30f · outbound

This paper cites Unlock the intermittent control ability of model free reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unlock the intermittent control ability of model free reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.069185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.627229Z digest=sha256:4e420468704103824e9a7c3005783df59576f4e1d6cec6fc8e79e9749ded4199

Observation 66a22057-2102-44d2-91ba-467c9ce178c7 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.060067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.676803Z digest=sha256:cf068ea71d42775c4e364831c6707767355dd50eddd01a78835c71292ea01249

Observation 777ddf33-8dd3-45f6-841b-b27fa644f0c7 · outbound

This paper cites W., and Parker-Holder, J.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., and Parker-Holder, J

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.049548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.718788Z digest=sha256:e96285f805abfd13ae077908d43c7c83f04c1e69c29da3432fb58a3b7e14bfe9

Observation 5a3c5aca-0cec-4191-a7ba-ccd0c65ce5ea · outbound

This paper cites Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.760568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.760568Z digest=sha256:6d460116abcfbd7409fc4a97c63771199b6447de999b169bd836b3ff21793bf1

Observation 914aae6e-5090-4352-802d-356c8f4d15ef · outbound

This paper cites Online reinforcement learning with uncertain episode lengths.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Online reinforcement learning with uncertain episode lengths

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.040234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.823709Z digest=sha256:52c15290bb370b18ee61fae31266054f3be51aa4274a3b521b0c726e422258dd

Observation 1bada11b-e190-455f-b112-7defdb9beb3d · outbound

This paper cites Tactical optimism and pessimism for deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Tactical optimism and pessimism for deep reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.029700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.877004Z digest=sha256:a3a4a7aec6c9ec7d2e01950eae479da41458d142400002d4b481ea2388626e02

Observation 0d1d9e27-5aed-4b82-a953-88d50f1a3aeb · outbound

This paper cites Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.919314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.919314Z digest=sha256:758d98e1457a3e35c39d5daf301837ba3f81a1ed66f8504e04b3c8ac88ccbc0b

Observation 070c5cfe-e979-4beb-bcc9-a55dbd64ec20 · outbound

This paper cites The primacy bias in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning The primacy bias in deep reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.016950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.968328Z digest=sha256:ef11aeda70702e7ca9615b540f947179b389abacc33fa4fbdfd7e6241c4b1860

Observation e8eb03d9-2105-45a2-bbaa-990e2397a778 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.004767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.017309Z digest=sha256:0a0c07912a269c609e31085271e4d966c1af230e0080915bb4b810a07243d116

Observation c0fc6932-682d-482a-bfc2-4920c266cc36 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.992729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.051990Z digest=sha256:641d240670d66b0e4ae9122e513486f1039dc8dfd7d1178024feb41b7288c7cf

Observation 8668d5d8-ef85-4d7b-8892-2c64c2f76568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.099844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.099844Z digest=sha256:39528c0a959497cd59be4d96ed53ea18b704c581292bcc1bb6624df8425c47a9

Observation 65a0b996-78b1-4ed1-a3ea-e92db6cdf013 · outbound

This paper cites Reinforcement Learning with Dynamic Boltzmann Softmax Updates.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Reinforcement Learning with Dynamic Boltzmann Softmax Updates

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.160250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.160250Z digest=sha256:8149e709616e0b6e31374765e88756c650cc236deb16fc7af8ab95f9b4457413

Observation 9660b867-6b68-419b-88d0-fd77ae2cf9f2 · outbound

This paper cites Softmax deep double deterministic policy gradients.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Softmax deep double deterministic policy gradients

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.981193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.192637Z digest=sha256:8590cf667d0b580f14dd2d5abceedcbd71b8b7436dfbe38d0090fc22b2722aa8

Observation fd8f489c-b566-440a-88c7-cf0253156be3 · outbound

This paper cites Regularized softmax deep multi-agent q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Regularized softmax deep multi-agent q-learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.970541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.230814Z digest=sha256:1289822962d7c396a5448123846dc95af542b57d283b0965fa21c253793fc27a

Observation 5c1641d6-a350-4942-84a0-4fc0f58aa313 · outbound

This paper cites Time limits in reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Time limits in reinforcement learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.958785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.321621Z digest=sha256:a7baae81f0c0692a9f65735682615c92ad2c5fd1a1ba76b805b968305a70f1b4

Observation a39b91ca-6935-48b4-a1b8-47d17ec83dd5 · outbound

This paper cites A., and Darrell, T.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning A., and Darrell, T

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.947216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.386837Z digest=sha256:4d0c30c441eb4d4856817fe8682f3841e82c98a1a4a2ddd6b5f450a0fb4b164a

Observation 2986c97b-8c8d-4f18-b235-2c43b089d4f4 · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.934747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.430420Z digest=sha256:accd60ee7124556107edea552ce4fbe55adf584d41944755fbd3ce84980eb604

Observation 42e3395c-bc25-4354-8cb0-42c32b25495c · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.921337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.493835Z digest=sha256:8837aa7d74eb8f9c95bbe2142409207d5e75243f00e2c901b49fe8b2b1f6c7c7

Observation ac7e3916-194d-49ab-9420-f5302e4f0da4 · outbound

This paper cites Optimistic Exploration even with a Pessimistic Initialisation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimistic Exploration even with a Pessimistic Initialisation

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.555362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.586211Z digest=sha256:d7e17de75d6fc324d516066f18536f8d103635a3c9673f2674d7950867ac8b90

Observation b7cee4a3-ece9-4414-85ac-3e7ec8f782a7 · outbound

This paper cites Prioritized Experience Replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritized Experience Replay

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.629886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.629886Z digest=sha256:4db36e68fd6db3f1540dd5d749f41e3bc863652de575ae5ff87c06632ed617c7

Observation 0e458a88-e4c5-4099-9924-1a5abfd97568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.908081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.697565Z digest=sha256:3a721dfe7788caaa507958b4176d7632aba328619527416ffcb669454646a060

Observation e9e3441b-e682-48b6-9f1b-4ea00bcb32b4 · outbound

This paper cites S., and Evci, U.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning S., and Evci, U

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.837429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.837429Z digest=sha256:00134864151d18e620a6b90767283a200ffc98a7fac2d4c05b5f1a6d6da95003

Observation 6a35abc8-8d00-4b36-893e-a38fb2906a11 · outbound

This paper cites Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.490510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.902691Z digest=sha256:e345f7a156eca99f1f97ef81763f2fee13f03383f1007935fa138161399b56d8

Observation f4ab1462-9de7-4f2a-8679-bc0723435502 · outbound

This paper cites Revisiting the softmax bellman operator: New benefits and new perspective.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting the softmax bellman operator: New benefits and new perspective

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.886519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.972733Z digest=sha256:b1e1e0641f219e4fcbb8b9ed113f7b07436f1cf9ad84076eb1964c2c2e226c9b

Observation d90db58a-4793-4d0b-9c9c-9099759ce176 · outbound

This paper cites Prioritizing samples in reinforcement learning with reducible loss.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritizing samples in reinforcement learning with reducible loss

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.876288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.084002Z digest=sha256:07d2a01424747e2bac3b6ab943fc67215f47e298901a9b0905d0673ec6ffd10e

Observation e6907165-b461-4f4a-ad31-8153b8e506b7 · outbound

This paper cites Safe Exploration by Solving Early Terminated MDP.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Safe Exploration by Solving Early Terminated MDP

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.325475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.197634Z digest=sha256:cc8468623d0671d29d5e6d57ceb073353dfe55f384117cd6ece0109f62eadb59

Observation 35d61335-ae2b-479d-abab-982bc9a399e0 · outbound

This paper cites sunk costs.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning sunk costs

Reference 56

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.587105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.278813Z digest=sha256:c0dfe9d2c03a6ba0057dce85ef64d838ac4b0a6d697d8fa52f403ace8fbcc351

Observation 364092c5-6388-4b31-afdc-14272e675950 · outbound

This paper cites DeepMind Control Suite.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DeepMind Control Suite

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.351761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.351761Z digest=sha256:640e035a16df8e75094e89acb1a61790434b7023fa93b211530d4143b8a1b3e0

Observation 6e83b4c8-5af9-4201-9939-05fa456612a6 · outbound

This paper cites Loss Functions and Metrics in Deep Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Loss Functions and Metrics in Deep Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.432291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.432291Z digest=sha256:4d3070bc8087256c13a77273c7d533687066dc34f9b11da17dc3855bf5957af4

Observation 388bbcb7-8741-43ce-a4a1-fe19799cd702 · outbound

This paper cites H., Meyers, E.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning H., Meyers, E

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.866665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.623704Z digest=sha256:18fa9bfb3738c010b4878628a95d26a47f7bd3b3b5343835c26378eadfcc704e

Observation 31640fed-96f4-4f9b-9cd6-58fb38c952ce · outbound

This paper cites Deep reinforcement learning with double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Deep reinforcement learning with double q-learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.857507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.697176Z digest=sha256:083fdde47485b83ec81527ed23cdc99bb448a9441d32a4babee506f793d3d353

Observation 58976586-1628-4cf5-a242-d26288add4d0 · outbound

This paper cites and Drake Jr, F.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning and Drake Jr, F

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.847097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.713465Z digest=sha256:ce40fe5f1b4f3eddf6f4fd45574cb61c71e6b58d3f345d08d814275b9d5cb3ff

Observation cc25d49c-21eb-4831-96f8-e5b9a1915127 · outbound

This paper cites Optimism in Reinforcement Learning with Generalized Linear Function Approximation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimism in Reinforcement Learning with Generalized Linear Function Approximation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.761065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.761065Z digest=sha256:35abfbf6f63e0f11f71d30f69de83f1c32c10dc5ce0a12bfa7555ec95a46f4c2

Observation 0f10747d-6b07-4af0-b75b-552092ec3032 · outbound

This paper cites DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.901414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.901414Z digest=sha256:3d3e564ae97b5fcf2a8fc28256a80f5c23950e9619849d1d436a4739867e70de

Observation c727a42e-9e6c-4af3-97cf-7fda9d49554c · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.985431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.985431Z digest=sha256:51f850adab26c8a223d4b2a4717d005077e22a0f772711bc2c6b243b9b814d23

Observation a72e8898-159d-41b9-a619-b35c1687e2dc · outbound

This paper cites Sample Efficient Deep Reinforcement Learning via Local Planning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sample Efficient Deep Reinforcement Learning via Local Planning

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.136783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.102945Z digest=sha256:d3a64a19c5a90f815670e707004fbc067b4fe5eb4b1a9b1ea99d3694272df874

Observation 0475043e-986d-4206-bde3-a48ea046d7fb · outbound

This paper cites Scaling Robot Learning with Semantically Imagined Experience.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Scaling Robot Learning with Semantically Imagined Experience

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.190932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.190932Z digest=sha256:37434412cbb442ba9ed02ed1991e3845c8e088eba377c9f60dfd84b074196c6a

Observation 0b575bae-4391-42df-9f0a-3a9a92463523 · outbound

This paper cites Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.835270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.287162Z digest=sha256:6b7f86c52e1d67958985d24e77fae8aea195398cf47ded891a9b7f0c8f787421

Observation ea620ca8-f979-4a01-ac43-e0762f4be893 · outbound

This paper cites write newline.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning write newline

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.400821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.400821Z digest=sha256:6bb742742bc6ddc6d8bc820d0db874ee6892a0207c5b8e32ce9c3c8d0fdb552d

Pith citing papers

No inbound Pith citation observations are available.