Pith. sign in

Paper Citation Record · LEDGER

Generalization and Regularization in DQN

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:1810.00123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1810.00123 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:46:07.253309Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T14:09:52.653481Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e08a654-056c-4ee5-807b-5b34b1eac042 · inbound

In Hindsight: A Smooth Reward for Steady Exploration cites this paper.

In Hindsight: A Smooth Reward for Steady Exploration Generalization and Regularization in DQN

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T17:41:05.708898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T17:38:45.913293Z digest=sha256:6088d34ab441e50aed3c80ab2ba138dd619102d2e0c1eab9d634e985ec1e9193

Observation a9fc6219-3c8c-4ac1-a151-91cc96beb823 · inbound

Generalizing from a few environments in safety-critical reinforcement learning cites this paper.

Generalizing from a few environments in safety-critical reinforcement learning Generalization and Regularization in DQN

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-25T11:00:40.168151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T10:57:45.056957Z digest=sha256:84e899215b57448fd7eb8572c77a4d3213bc43973d8b74e6a2830f848b4f3740

Observation 30d6457c-ad63-4e72-b30f-4c1d74fdf07d · inbound

Reasoning and Generalization in RL: A Tool Use Perspective cites this paper.

Reasoning and Generalization in RL: A Tool Use Perspective Generalization and Regularization in DQN

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T09:20:34.443759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T09:20:04.036102Z digest=sha256:2ed0bfdc95aa16956eaf58625e6be45852ac46a0247122ec68fc8677c5135ef6

Observation ca5ac14d-9989-4a86-8d63-531a5dc3b7e5 · inbound

Scaling Laws for Reward Model Overoptimization cites this paper.

Scaling Laws for Reward Model Overoptimization Generalization and Regularization in DQN

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:04:53.245501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T09:04:53.129737Z digest=sha256:713299731955aaf7b55741f41124a6e6a0720c4697dba7cb5dc276113f1a09cb

Observation 6cded5cc-c288-4cbe-b46c-e7aeea35f7a0 · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey Generalization and Regularization in DQN

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:50.692503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:fcdcc98bdfc8ce5623551264594a45550855059365d48df875c0726e53df6ad3

Observation 63b110f6-35da-4ff4-9f02-171cea97c91b · inbound

Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning cites this paper.

Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T17:46:07.253309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:46:07.253309Z digest=sha256:d0469b7c41bffde3fc7a58b649b789e91ea83f1503ee137bbc4d9a540681a604

Observation ec076800-faf5-4fb0-86b8-f4e561d1f171 · inbound

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing cites this paper.

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing Generalization and Regularization in DQN

Reference 1967

Resolution
unresolved
no resolver link, observed 2026-08-09T11:22:28.868513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:22:28.868513Z digest=sha256:b1994d4262481e122786b5370ac26b5b7bf0cac4fa4b953be2c5d214f3994bcb

Observation 91e7bc31-a76d-4b71-9efb-2ad424663a00 · inbound

Hierarchical Successor Representation for Robust Transfer cites this paper.

Hierarchical Successor Representation for Robust Transfer Generalization and Regularization in DQN

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-02T23:46:56.405225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:46:56.405225Z digest=sha256:53b82829a7be87dd64990d60ffdf312121ddbf95c3bec0ee0742a5ae48a0d8dc

Observation 13e803c1-d8d9-416a-8132-ae5706b5c6dd · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.720409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:72a9ac22c6e3e47c7a582a4d71176a2fcc8646f425ececb78c27fffc05fd048d

Observation ab828019-ff6e-41fd-8278-6b2e670873be · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.822838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:82ad4e910c474776e2661ac404c87d2589bac55a0fb1cea2b48affc89bfefe7b

Observation 0d9ae354-4fd0-407f-88ab-d791e2642e16 · inbound

Bridging Performance and Generalization in Reinforcement Learning for Agile Flight cites this paper.

Bridging Performance and Generalization in Reinforcement Learning for Agile Flight Generalization and Regularization in DQN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:09:52.655148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T04:33:20.981811Z digest=sha256:e390f51ba1cd782d9a0ee2139edb5c88767b92e56922dec0966b327903bc67eb