Pith. sign in

Paper Citation Record · LEDGER

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning

As of 17 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2411.08360.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08360 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:48:16.963053Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact1
  • verified fuzzy28
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 25ddcad1-4e33-4998-b32e-9fd794e2fa47 · outbound

This paper cites Reinforcement learning: An introduction.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Reinforcement learning: An introduction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.793163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.793163Z digest=sha256:5124b8958c3210870e885962a1aa15cbc8b7540b9ba341db82c1a64a6d589908

Observation e33c4ae6-81d7-4c58-90bd-503960edf392 · outbound

This paper cites Reliable adaptive recoding for batched network coding with burst-noise channels.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Reliable adaptive recoding for batched network coding with burst-noise channels

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.614969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.798009Z digest=sha256:58e95aa7e9b3826f6203d2f35e253bd0c08e374b0bb3df2f646261b884766254

Observation df5fc6d5-2f8c-4203-ad21-3b5818c1a4a2 · outbound

This paper cites Markov decision processes with applications in wireless sensor networks: A survey.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Markov decision processes with applications in wireless sensor networks: A survey

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.601632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.802596Z digest=sha256:0b1f24472017e28888cc60e3c7b96cd1a454a0bea6f75040885806f8d2d0bcc1

Observation 8aa9d975-95bb-486c-84ea-e591686ef37a · outbound

This paper cites Q-learning algorithms: A comprehensive classification and appli- cations.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Q-learning algorithms: A comprehensive classification and appli- cations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.589829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.807163Z digest=sha256:70e45920b2be8f5a7d3e2360f1cfb9278be0003f72d4ddc7c85d2b50732b34d4

Observation 5f436d24-5220-4a6a-81ca-e70262b8f621 · outbound

This paper cites Q-learning: Theory and applications.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Q-learning: Theory and applications

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.577270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.811366Z digest=sha256:b8233222958366afde1d4e269a9b89912816b460d082752a6772e00162146794

Observation 9f69c23c-ae2d-45c6-a9ae-8c326e57506a · outbound

This paper cites Double Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Double Q-learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.564179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.815360Z digest=sha256:23112776b8e130a5f74c655168a64f4a719d52a65f0928936aa67a2d95ccd1a9

Observation 14c6601e-c24c-4952-8378-c58e0e32367b · outbound

This paper cites Ensemble bootstrapping for Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Ensemble bootstrapping for Q-learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.548708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.819736Z digest=sha256:16daad66a8fee86976da23926991bfb4e1353c412e87c62e1da26a91fca32a83

Observation a2fbc0ca-5be0-4711-98bd-5ba95cd1fadf · outbound

This paper cites Maxmin Q-learning: Controlling the Estimation Bias of Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Maxmin Q-learning: Controlling the Estimation Bias of Q-learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.823437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.823437Z digest=sha256:e20984496206ee18396ae4ec84f3dec791d42677a864943a61f9bb5472c8ac19

Observation 788b4e98-d8e9-49e1-9cde-98ff85b29814 · outbound

This paper cites Speedy Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Speedy Q-learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.534597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.827628Z digest=sha256:85d77dc5ffb5b202e464880e50d4eb12ab8b8938fa18178c303d8c79b831870a

Observation 437f5186-876e-428d-ad25-e206beb3b5ae · outbound

This paper cites Pac model-free reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Pac model-free reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.503165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.831949Z digest=sha256:76edb6c4b7ea8ec977197e8f4dd8803ae82b296cafa51cc4885214607bd731ae

Observation fdb3a6f6-4545-41b9-bcbd-6b76f079f28b · outbound

This paper cites Neural fitted Q iteration–first experiences with a data efficient neural reinforcement learning method.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Neural fitted Q iteration–first experiences with a data efficient neural reinforcement learning method

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.490087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.835771Z digest=sha256:b62635405912b08170a3961925364520528f8d36996b6ce7c51da6f83dd61442

Observation e8529386-077f-48d3-aff5-270c99a98bdd · outbound

This paper cites Deep exploration via bootstrapped dqn.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Deep exploration via bootstrapped dqn

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.839747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.839747Z digest=sha256:1d13201347f5d19655dd8bd88896f04e1bdc5b1989dd97cd84387f0291206807

Observation 437aba3e-ae67-4b77-9ca5-84988a9186ab · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Playing Atari with Deep Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.843382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.843382Z digest=sha256:6a59e143e9230321478736c367a5bda8b754659617708faceaa999d8c1a3157b

Observation 93088779-e9d7-477c-9357-8f8d585e4d46 · outbound

This paper cites Q-learning with linear function approximation.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Q-learning with linear function approximation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.469419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.847237Z digest=sha256:13bad5aa56efbe6d3dd86edbdf483d3f13e16de152b601bb483eb8a64050a61b

Observation c4b897ad-5a0e-468c-a703-3d727afdbec3 · outbound

This paper cites Sample complexity of reinforcement learning using linearly combined model ensembles.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Sample complexity of reinforcement learning using linearly combined model ensembles

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.438526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.854287Z digest=sha256:1c3ccde669ac41ea5608dfd9579ee6c6fa02671b870164ef68d0229e950e1e79

Observation 2cdac10e-b52c-4cb3-aeb4-900e4b5f0a31 · outbound

This paper cites Model-Ensemble Trust-Region Policy Optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Model-Ensemble Trust-Region Policy Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.858060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.858060Z digest=sha256:017f51b7bfcca9aa6d66c7b67f7b6efd9166f940fcec61cecc688f43d899710e

Observation 9fc40cfa-629d-46e3-adf7-3a93dd311a94 · outbound

This paper cites Deep reinforcement learning in a handful of trials using prob- abilistic dynamics models.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Deep reinforcement learning in a handful of trials using prob- abilistic dynamics models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.862136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.862136Z digest=sha256:d9dd8eabf843c7f4346d3e710576ae425c1f9dc5bc9dc32d372af3e2e379ca76

Observation bf293276-88c3-4155-ad35-927b6832fecf · outbound

This paper cites Asynchronous methods for deep reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Asynchronous methods for deep reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.865881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.865881Z digest=sha256:de59dba42beaf4da08d7d90f2e7a46f7a13cf9dd6dc470dda9942fb2de502d91

Observation 35e61a55-621d-4e71-9e1b-795e9e89febf · outbound

This paper cites Ensemble link learning for large state space multiple access communications.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Ensemble link learning for large state space multiple access communications

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.392417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.870066Z digest=sha256:ec1525e3d9379c929da76eca8d7dc65266260a6f1a2cc382470607871d781531

Observation f911ecda-c405-4234-bd60-bd8ea2784f04 · outbound

This paper cites Ensemble graph Q-learning for large scale networks.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Ensemble graph Q-learning for large scale networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.377898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.873712Z digest=sha256:c9de41934de338b99256abd9877f92c5c5e90053085449ac74120503221da35c

Observation f1f5829f-c46c-415a-9af1-6c27bc24101b · outbound

This paper cites Multi-timescale ensemble q-learning for markov decision process policy optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Multi-timescale ensemble q-learning for markov decision process policy optimization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.363970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.877389Z digest=sha256:ad644efd999643faf31772c2a97c9ee3c4fcb8b665b865bd7e360f38b179c09d

Observation c4653aeb-3949-4bef-acfc-97fae01d8eb2 · outbound

This paper cites Leveraging digital cousins for ensemble q-learning in large-scale wireless networks.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Leveraging digital cousins for ensemble q-learning in large-scale wireless networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.348408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.882138Z digest=sha256:32e4c6c477f9aa22c21d14ddef2cec88b3c580307c39b69a7d320575ef771d18

Observation f7c8391b-7c76-480b-a1b7-38c697c05aa4 · outbound

This paper cites A novel ensemble q-learning algorithm for policy optimization in large-scale networks.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning A novel ensemble q-learning algorithm for policy optimization in large-scale networks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.335126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.885810Z digest=sha256:43af2c78a4ef0bee6ff4c10b40e7c573cbb4927978f6305201b3770934cf8901

Observation 94fe94b0-4fbe-4b16-9f65-7eb442b8d7c1 · outbound

This paper cites Link analysis for solving multiple- access mdps with large state spaces.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Link analysis for solving multiple- access mdps with large state spaces

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.322383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.889556Z digest=sha256:5cecdcd65037b30589b073ae28eeed2ee3bb8b1d5d44eecd49b21f5c80551fc2

Observation d3692961-f7ac-4542-aa87-34bbb7cef735 · outbound

This paper cites Is q-learning provably efficient? Advances in neural information processing systems, 31, 2018.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Is q-learning provably efficient? Advances in neural information processing systems, 31, 2018

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.893407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.893407Z digest=sha256:382a4117ba0b5c30d81e0fa7edd16951b4053d9716a793dfb64b835aac5bc8ab

Observation 8d6795f2-3ec8-4547-b960-5b58ae3ea6ff · outbound

This paper cites Online robust reinforcement learning with model uncertainty.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Online robust reinforcement learning with model uncertainty

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.897232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.897232Z digest=sha256:b234f3f4f6f88e9f8cf4fa5a1f5d923dc0b51bf2ac80eced55789ebfcb333c82

Observation da4cd2ce-b9f7-437a-a4a1-ddeb4688a324 · outbound

This paper cites Conser- vative q-learning for offline reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Conser- vative q-learning for offline reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.285828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.901220Z digest=sha256:d82aa8b43dca40eacbce2b5b8bdf035cbff3f33f8536586b8394607e016e7b0b

Observation f0ea9ff9-5f10-44ff-b982-5d7ca4bfe099 · outbound

This paper cites An optimistic perspective on offline reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning An optimistic perspective on offline reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.273193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.905076Z digest=sha256:c78bcb6d616ead965d749d506db7526bcaea9f67c39402ab12f6c8753310fd53

Observation 2b6f2766-2cab-49cd-adc5-e74516f10fdb · outbound

This paper cites Leveraging offline data in online reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Leveraging offline data in online reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.260691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.908840Z digest=sha256:f7499afb9190784ce6044959ca48ab75adbdb3ed2fb2744de2efdf6f47b8a01d

Observation 1b4cfa2c-1a0c-4510-bd0b-10d85210e379 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.912632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.912632Z digest=sha256:8548f9eee09ad228477654ff8690a99803e0d9ff4d60e58e8d0121f0b59a630e

Observation a7cf1705-e583-4563-811c-3aa513f289a5 · outbound

This paper cites Policy finetuning: Bridging sample-efficient offline and online reinforce- ment learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Policy finetuning: Bridging sample-efficient offline and online reinforce- ment learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.247073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.916639Z digest=sha256:55129a7bd5ff0cfdd29bdbeac4fd5e5849757e413ff30a8ee1c86c48317c95f0

Observation 2e6d1209-d8cc-4956-aa2c-08bc1c594bdc · outbound

This paper cites Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.920591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.920591Z digest=sha256:eb3d43fd205d30c1079ea1c20061e2734b03e34bc0cb7bdd1959c11bf3500e82

Observation af046d29-91fc-40f3-9395-0c54ed904789 · outbound

This paper cites Com- paring exploration strategies for q-learning in random stochastic mazes.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Com- paring exploration strategies for q-learning in random stochastic mazes

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.230930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.924650Z digest=sha256:ea33eada8b315b764739060ba7225fdb52b4772936cff0f966a1316aa63e1afb

Observation b10298f2-4e93-40ef-864d-5fcf11b9e2d3 · outbound

This paper cites The Role of Coverage in Online Reinforcement Learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning The Role of Coverage in Online Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.928379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.928379Z digest=sha256:20b44be7cc9fec74664320e143e42b34060d80ea537790dd38a639e91d507342

Observation 17fc7772-cfc9-4e89-83f4-c044a97f4ba1 · outbound

This paper cites Pessimistic Model-based Offline Reinforcement Learning under Partial Coverage.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Pessimistic Model-based Offline Reinforcement Learning under Partial Coverage

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.932746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.932746Z digest=sha256:f73635697fced123a313b79b07331d385eb2cde85b198fcac6f6167b77eb3b14

Observation 0994f7b3-a7db-4031-a407-46709bcebc5f · outbound

This paper cites What can online reinforcement learning with function approximation benefit from general coverage conditions? In International Conference on Machine Learning, pages 22063–22091.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning What can online reinforcement learning with function approximation benefit from general coverage conditions? In International Conference on Machine Learning, pages 22063–22091

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.211572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.936660Z digest=sha256:e16d8eecadc5577441fd8d99c0391b3c4f488ff59ab6301f472c87afbd90af4f

Observation 14212655-35fd-4a34-ac2e-4f20997ff5ef · outbound

This paper cites When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.196846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.940435Z digest=sha256:35ca2850f2130fbbab530545d59322417acbfc8099093dfff3c75d72e405ffb7

Observation 4a951e64-b95d-4ded-a363-bd8dd1f1a35d · outbound

This paper cites Coverage analysis of multi- environment q-learning algorithms for wireless network optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Coverage analysis of multi- environment q-learning algorithms for wireless network optimization

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.174530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.944097Z digest=sha256:7b843308b02fa12c2758d1147162236b6c3c54865e94b440532e590b2cf605cf

Observation 62ba41a5-6357-4d1d-b602-938ee6d2ab46 · outbound

This paper cites Convergence of Q-learning: A simple proof.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Convergence of Q-learning: A simple proof

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.160552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.947548Z digest=sha256:5a1bf5117c06e74791d054a8b612368651afb31b1f33df8a8e5748ac9f99eb74

Observation 1afabcb6-37a8-4099-9fc0-3406415ba7bf · outbound

This paper cites Issues in using function ap- proximation for reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Issues in using function ap- proximation for reinforcement learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.140862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.951314Z digest=sha256:a84694b927301b7100693cc9f199fdf678119484bde9be44adbc71dda95cf90a

Observation 244acb39-d088-4af2-aa02-2e23dee15268 · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.955046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.955046Z digest=sha256:b5091ac941c1dde4fb6de85359186ebc5e8267799d9ab79394541fe20e54d668

Observation eadff9d3-985d-40c0-8e74-1cf2992aabf5 · outbound

This paper cites The virtues of laziness in model-based rl: A unified objective and algorithms.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning The virtues of laziness in model-based rl: A unified objective and algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.959107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.959107Z digest=sha256:d20d87a5b77135e74c3ab03a36f86fbd09f98ecefe18c65a9393d16fd542c48f

Observation e72cc9b0-35ce-41e7-9d18-6532afba8694 · outbound

This paper cites A Multi-Agent Multi-Environment Mixed Q-Learning for Partially Decentralized Wireless Network Optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning A Multi-Agent Multi-Environment Mixed Q-Learning for Partially Decentralized Wireless Network Optimization

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T21:48:17.003329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.963053Z digest=sha256:7cf3f4285a604fe9307e73f2279d798b33e9282a0043fdc1183849a65ebbf12b

Observation 4dcbefa6-4607-4d8d-8b7f-c1537b30145f · outbound

This paper cites Springer, 2007.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Springer, 2007

Reference 2007

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.455306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T21:48:16.850593Z digest=sha256:bfdd2882b1a9dfa0f652c2fb49871f2cbaf0b558b409001b5a254eda2bbf9503

Pith citing papers

No inbound Pith citation observations are available.