Pith. sign in

Paper Citation Record · LEDGER

Bellman operator convergence enhancements in reinforcement learning algorithms

As of 11 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 2 inbound Pith citation observations for arXiv:2505.14564.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14564 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:23.070508Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T19:15:30.759880Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T19:20:31.176550Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact3
  • verified fuzzy13
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 97971769-4e0c-44b1-aa2e-c90c255db885 · outbound

This paper cites Accessed on 13/03/2024.

Bellman operator convergence enhancements in reinforcement learning algorithms Accessed on 13/03/2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:26.040351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:21.194203Z digest=sha256:229cf8bd6b4f2699f8578dd0f3dd26375760d7e1d638ee1095c2ca14c8e7e543

Observation 88a77697-2960-4bfc-a9da-32b850b2e55e · outbound

This paper cites An alternative softmax operator for reinforcement learning.

Bellman operator convergence enhancements in reinforcement learning algorithms An alternative softmax operator for reinforcement learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:21.276599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:21.276599Z digest=sha256:92965e54d1edf50ca7b06555ecc9c0ee9b4f88659b50747384ce9fa1e469a65b

Observation 22af8a4b-41b4-401c-b211-b638843799b6 · outbound

This paper cites Lipschitz Continuity in Model-based Reinforcement Learning.

Bellman operator convergence enhancements in reinforcement learning algorithms Lipschitz Continuity in Model-based Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:21.402923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:21.402923Z digest=sha256:613d2693b7e1c2cdc6b27d42828b32655fb2e873010d47c9f23d520c0c908d8d

Observation 30701fa9-700d-44a8-a382-4029391ac28f · outbound

This paper cites Speedy q-learning.

Bellman operator convergence enhancements in reinforcement learning algorithms Speedy q-learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.931030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:21.519597Z digest=sha256:842d4242e6607cda3c7df38ba5d60b818996b10af08557b23b3061887c8a6031

Observation c70850de-13a6-4feb-9af8-4da18a2eeeb7 · outbound

This paper cites Neuronlike adaptive elements that can solve difficult learning control problems.IEEE transactions on systems, man, and cybernetics, (5):834–846, 1983.

Bellman operator convergence enhancements in reinforcement learning algorithms Neuronlike adaptive elements that can solve difficult learning control problems.IEEE transactions on systems, man, and cybernetics, (5):834–846, 1983

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.847166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:21.610842Z digest=sha256:2b1a6101e4b089f8f8858c760fa6e16d51f2a0e9df874df19ba6747d17749a11

Observation 385bb917-34a3-4712-8395-c5d0f5b6d3d3 · outbound

This paper cites Increasing the action gap: New operators for reinforcement learning.

Bellman operator convergence enhancements in reinforcement learning algorithms Increasing the action gap: New operators for reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.694636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:21.749774Z digest=sha256:9d70c1c5335c15963b77ead2a23e972f6aaaac42b52447dc5ef8a1d131bd54f6

Observation 5dc4b47a-a18f-475e-a079-cbee99a440f0 · outbound

This paper cites Q-learning and enhanced policy iteration in discounted dynamic programming.

Bellman operator convergence enhancements in reinforcement learning algorithms Q-learning and enhanced policy iteration in discounted dynamic programming

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.575426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:21.837142Z digest=sha256:98728474215fb5d3841d72087d173df92d686aabbc7fabee06937e35c828b3dc

Observation ee427491-64a4-4073-bbf6-1bbd0e756822 · outbound

This paper cites The Value Function Polytope in Reinforcement Learning.

Bellman operator convergence enhancements in reinforcement learning algorithms The Value Function Polytope in Reinforcement Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:36:23.995209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:21.874040Z digest=sha256:59fb23a471eea484aca59c8e96e0adda88248e00517367c95cef42372068d491

Observation 3396a85f-3cab-412a-ba59-edf573264ac2 · outbound

This paper cites Addison-Wesley Professional, 2019.

Bellman operator convergence enhancements in reinforcement learning algorithms Addison-Wesley Professional, 2019

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.443561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.015595Z digest=sha256:d4f8cf2e67ced0576f461f35cf9ca5373d2e05dc58a1e180e8a33245effcbfa5

Observation 077553d5-df00-491a-93ba-764231b43acd · outbound

This paper cites Topological Foundations of Reinforcement Learning.

Bellman operator convergence enhancements in reinforcement learning algorithms Topological Foundations of Reinforcement Learning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:36:23.723956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.146406Z digest=sha256:6a1d4be6f7f6c5bfd69237bdd95d8f5b9eea55f5b8f9fb937e658488208034da

Observation 94e157b0-8bd1-4747-a679-eb98229684e7 · outbound

This paper cites Reinforcement learning essay (aims-cameroon).

Bellman operator convergence enhancements in reinforcement learning algorithms Reinforcement learning essay (aims-cameroon)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.283575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.248741Z digest=sha256:d4f43bbe6ff63208f820dd181987e7bccb57b68ca3cd33361e0516a79df2a5a9

Observation 12ddeb6d-a788-43bb-b167-c188ddef1d34 · outbound

This paper cites Metrics and continuity in reinforcement learning.

Bellman operator convergence enhancements in reinforcement learning algorithms Metrics and continuity in reinforcement learning

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:36:23.456679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.336319Z digest=sha256:95bf0612b47ac253a80bf3ae7441a497777b7f2f52fa7b1da441dfe7ce84f532

Observation 47ff5210-c516-4a95-be5e-bd0fa24ae701 · outbound

This paper cites Markov decision processes and dynamic programming, 2013.

Bellman operator convergence enhancements in reinforcement learning algorithms Markov decision processes and dynamic programming, 2013

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:25.162287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.421040Z digest=sha256:db4b0a9d8f3e4215c1f6f5554f959aab87d5384e21f005cef7350b2ff7c033ed

Observation 3e58590e-34a4-433b-8663-793648d0ee5d · outbound

This paper cites A General Family of Robust Stochastic Operators for Reinforcement Learning.

Bellman operator convergence enhancements in reinforcement learning algorithms A General Family of Robust Stochastic Operators for Reinforcement Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:36:23.221020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.530164Z digest=sha256:34ea6ef30b81b60d68da539721078958e63ee7c5cf07af0b0f5b809a3b7a7056

Observation 83833abb-6141-4a9d-ac0f-03679beb8b2e · outbound

This paper cites Efficient memory-based learning for robot control.

Bellman operator convergence enhancements in reinforcement learning algorithms Efficient memory-based learning for robot control

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:24.998711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.608027Z digest=sha256:1c8100249bc41ad7285b36f27d645876fbe0f94af897f41d10c2b1eebad71134

Observation a2041a33-4389-44b0-b9f4-c833db4bdcbf · outbound

This paper cites John Wiley & Sons, 2013.

Bellman operator convergence enhancements in reinforcement learning algorithms John Wiley & Sons, 2013

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:24.857823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.685167Z digest=sha256:a12d8cdf4b3b00139f55af7740908f5ee27accf9014a637567a7686adad3eb17

Observation ef987062-9d92-4e6a-bf8b-2526c97377cb · outbound

This paper cites Efficient Model-free Reinforcement Learning in Metric Spaces.

Bellman operator convergence enhancements in reinforcement learning algorithms Efficient Model-free Reinforcement Learning in Metric Spaces

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:22.717341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:22.717341Z digest=sha256:f11c9e16f3137f164e2183af3cb1c6ae6bbd49d4bb700cdb0d59ed30246e772b

Observation 604ee969-6e6f-4f26-b1d7-2859d4bf060b · outbound

This paper cites Generalization in reinforcement learning: Successful examples using sparse coarse coding.Advances in neural information processing systems, 8, 1995.

Bellman operator convergence enhancements in reinforcement learning algorithms Generalization in reinforcement learning: Successful examples using sparse coarse coding.Advances in neural information processing systems, 8, 1995

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:22.778518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:22.778518Z digest=sha256:13029520cc8697015364c4d301b10b9e2fcc47e898d1b43a546ad56de27ebb8a

Observation e5dc5b02-45a1-42f9-9128-f1da5c38c23c · outbound

This paper cites MIT press, 2018.

Bellman operator convergence enhancements in reinforcement learning algorithms MIT press, 2018

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:22.839696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:22.839696Z digest=sha256:8f6d09c5aaf756d57147ad23119c4c16716e352bce8de483532ef4e3ab6b22fc

Observation 4633efd3-8bb1-47ae-95f3-7945135c3b9e · outbound

This paper cites Nova Science Publishers, New York, 2021.

Bellman operator convergence enhancements in reinforcement learning algorithms Nova Science Publishers, New York, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:24.689116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.874997Z digest=sha256:ca1f3519ce9c4d6ba8379d4b8e64047d26287414908c82dd3e66345bd094cc4c

Observation 024b3833-47fe-415b-b7c1-5109dbabce50 · outbound

This paper cites Github : Basic Reinforcement Learning.

Bellman operator convergence enhancements in reinforcement learning algorithms Github : Basic Reinforcement Learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:24.580346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:22.981197Z digest=sha256:9ef4c63b407fda7db1e89cef41981f3687a9c224fdf0ebde56b2cb085c7b2d50

Observation 31244681-40c7-4d4d-af3f-5c18478d4b53 · outbound

This paper cites Learning from delayed rewards.

Bellman operator convergence enhancements in reinforcement learning algorithms Learning from delayed rewards

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:24.241206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T15:36:23.070508Z digest=sha256:14d205c574d948bac8b2384eeb2a715476f906fca53a34cdcc8bf5d382c573a1

Pith citing papers

Observation 5e064f49-fe9c-44bf-a386-8cd8c3911954 · inbound

Carbon-Aware Intrusion Detection: A Comparative Study of Supervised and Unsupervised DRL for Sustainable IoT Edge Gateways cites this paper.

Carbon-Aware Intrusion Detection: A Comparative Study of Supervised and Unsupervised DRL for Sustainable IoT Edge Gateways Bellman operator convergence enhancements in reinforcement learning algorithms

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:20:31.178252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T19:15:30.759880Z digest=sha256:5f8b92067edf1a4e909ac889a2cb73e52455b06139b78b080dfc6159ed0958c4

Observation 92b1a724-e897-400f-b5a6-3aec6e20797d · inbound

TabQL: In-Context Q-Learning with Tabular Foundation Models cites this paper.

TabQL: In-Context Q-Learning with Tabular Foundation Models Bellman operator convergence enhancements in reinforcement learning algorithms

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:38:16.878324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T12:34:58.734670Z digest=sha256:401942541f9a5a5f2a03a6f0a19e54f5f297ac1ad268d66f9ae63ae54cc1ef3c