Pith. sign in

Paper Citation Record · LEDGER

Generalization and Regularization in DQN

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:1810.00123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1810.00123 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:48:46.569656Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T14:09:52.653481Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e08a654-056c-4ee5-807b-5b34b1eac042 · inbound

In Hindsight: A Smooth Reward for Steady Exploration cites this paper.

In Hindsight: A Smooth Reward for Steady Exploration Generalization and Regularization in DQN

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T17:41:05.708898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T17:38:45.913293Z digest=sha256:857143feae29d01eb12e5b1d8439233220467c0fb4f6c94ddad590eb15fb4562

Observation a9fc6219-3c8c-4ac1-a151-91cc96beb823 · inbound

Generalizing from a few environments in safety-critical reinforcement learning cites this paper.

Generalizing from a few environments in safety-critical reinforcement learning Generalization and Regularization in DQN

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-25T11:00:40.168151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T10:57:45.056957Z digest=sha256:df23dfe5ac8f2f2e7e4de2662311304fcc7dd72155e76b08bfad9158b436eb7c

Observation 30d6457c-ad63-4e72-b30f-4c1d74fdf07d · inbound

Reasoning and Generalization in RL: A Tool Use Perspective cites this paper.

Reasoning and Generalization in RL: A Tool Use Perspective Generalization and Regularization in DQN

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T09:20:34.443759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T09:20:04.036102Z digest=sha256:390cf773bf995b00218b4e8f358d1e66fcfcdc779dd924263e190171ca8edc96

Observation ca5ac14d-9989-4a86-8d63-531a5dc3b7e5 · inbound

Scaling Laws for Reward Model Overoptimization cites this paper.

Scaling Laws for Reward Model Overoptimization Generalization and Regularization in DQN

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:04:53.245501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T09:04:53.129737Z digest=sha256:9e7bdc6611f6cb8c1d808a163cd3fea5960846d8c3172bddce20e6469e3f0f2d

Observation 6cded5cc-c288-4cbe-b46c-e7aeea35f7a0 · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey Generalization and Regularization in DQN

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:50.692503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:8d246e94925228dc3e85d6d074f1a4a88827da77fafbc9c4ffeb81a874e2aecd

Observation 9a183f81-0082-471d-beb2-14c07b8694a8 · inbound

A Research Agenda for Usability and Generalisation in Reinforcement Learning cites this paper.

A Research Agenda for Usability and Generalisation in Reinforcement Learning Generalization and Regularization in DQN

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T05:58:46.294480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:58:46.294480Z digest=sha256:a579cc4780701343a5546ad5c5bcca1ba00d33201af63cf50abaaacc2f810ef6

Observation 63b110f6-35da-4ff4-9f02-171cea97c91b · inbound

Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning cites this paper.

Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T17:46:07.253309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:46:07.253309Z digest=sha256:f0f530a349228ba9bac79ee5610e32aff93954c55612b948ed8690bee79e80be

Observation ec076800-faf5-4fb0-86b8-f4e561d1f171 · inbound

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing cites this paper.

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing Generalization and Regularization in DQN

Reference 1967

Resolution
unresolved
no resolver link, observed 2026-08-09T11:22:28.868513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:22:28.868513Z digest=sha256:b1994d4262481e122786b5370ac26b5b7bf0cac4fa4b953be2c5d214f3994bcb

Observation 91e7bc31-a76d-4b71-9efb-2ad424663a00 · inbound

Hierarchical Successor Representation for Robust Transfer cites this paper.

Hierarchical Successor Representation for Robust Transfer Generalization and Regularization in DQN

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-02T23:46:56.405225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:46:56.405225Z digest=sha256:53b82829a7be87dd64990d60ffdf312121ddbf95c3bec0ee0742a5ae48a0d8dc

Observation 13e803c1-d8d9-416a-8132-ae5706b5c6dd · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.720409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:2ea8b4a210c31875c3e4154a9c983fd1e465aa68425d945cb0ddf4f994ddbd8d

Observation ab828019-ff6e-41fd-8278-6b2e670873be · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.822838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:5d71a500ce1d0338aa66080e7a45f55042b936537b5f8114d7791f58303beed6

Observation 0d9ae354-4fd0-407f-88ab-d791e2642e16 · inbound

Bridging Performance and Generalization in Reinforcement Learning for Agile Flight cites this paper.

Bridging Performance and Generalization in Reinforcement Learning for Agile Flight Generalization and Regularization in DQN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:09:52.655148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T04:33:20.981811Z digest=sha256:be8d0a2e4dd7b15af7a430c4ebaa1b12e2e7b950350cd406f80bfce244549fc3

Observation 2a199256-a862-47db-81e5-fb823b2538c1 · inbound

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control cites this paper.

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control Generalization and Regularization in DQN

Reference 236

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:46.569656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:46.569656Z digest=sha256:ddb6bd07da34a2372b06f46140aef101553bbf405f517da603afc307d2526feb