Pith. sign in

Paper Citation Record · LEDGER

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning

As of 17 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2411.08360.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08360 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:48:16.963053Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact1
  • verified fuzzy28
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 25ddcad1-4e33-4998-b32e-9fd794e2fa47 · outbound

This paper cites Reinforcement learning: An introduction.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Reinforcement learning: An introduction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.793163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.793163Z digest=sha256:5124b8958c3210870e885962a1aa15cbc8b7540b9ba341db82c1a64a6d589908

Observation e33c4ae6-81d7-4c58-90bd-503960edf392 · outbound

This paper cites Reliable adaptive recoding for batched network coding with burst-noise channels.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Reliable adaptive recoding for batched network coding with burst-noise channels

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.614969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.798009Z digest=sha256:10d05aee3856338a822ac632516428f2a4b4d8ebabeaf1a5cd4ec28621face20

Observation df5fc6d5-2f8c-4203-ad21-3b5818c1a4a2 · outbound

This paper cites Markov decision processes with applications in wireless sensor networks: A survey.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Markov decision processes with applications in wireless sensor networks: A survey

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.601632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.802596Z digest=sha256:cbc02465e69f0bad29cb8dcb1ec10226c8bb10b9abe14bb4e450356c5178567d

Observation 8aa9d975-95bb-486c-84ea-e591686ef37a · outbound

This paper cites Q-learning algorithms: A comprehensive classification and appli- cations.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Q-learning algorithms: A comprehensive classification and appli- cations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.589829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.807163Z digest=sha256:b6fcee8278667f92b6f33706d440011a8591fcdb199b0a43bfcc65ff20fb6e14

Observation 5f436d24-5220-4a6a-81ca-e70262b8f621 · outbound

This paper cites Q-learning: Theory and applications.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Q-learning: Theory and applications

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.577270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.811366Z digest=sha256:c22230d18feb24cd4cbab77f75f39d28bc3b9c925f35693e0d766da830c657d7

Observation 9f69c23c-ae2d-45c6-a9ae-8c326e57506a · outbound

This paper cites Double Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Double Q-learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.564179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.815360Z digest=sha256:fe6880050d09920f8d03a91be3ea5fb9e18176f1a51da2fef977f44ee359cbf6

Observation 14c6601e-c24c-4952-8378-c58e0e32367b · outbound

This paper cites Ensemble bootstrapping for Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Ensemble bootstrapping for Q-learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.548708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.819736Z digest=sha256:0452df0e9c352ed6c17b591290ab49d8a07b98d7163c8b9041e17bdaed494e2b

Observation a2fbc0ca-5be0-4711-98bd-5ba95cd1fadf · outbound

This paper cites Maxmin Q-learning: Controlling the Estimation Bias of Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Maxmin Q-learning: Controlling the Estimation Bias of Q-learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.823437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.823437Z digest=sha256:e20984496206ee18396ae4ec84f3dec791d42677a864943a61f9bb5472c8ac19

Observation 788b4e98-d8e9-49e1-9cde-98ff85b29814 · outbound

This paper cites Speedy Q-learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Speedy Q-learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.534597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.827628Z digest=sha256:5f399f5770d691fe7e27262c1371dd4393d722ea524ae83f98301a9c4d1e65d7

Observation 437f5186-876e-428d-ad25-e206beb3b5ae · outbound

This paper cites Pac model-free reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Pac model-free reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.503165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.831949Z digest=sha256:7583a456704a65b802959dcabcf7c08fc219344349ba7b53b7312c518219060c

Observation fdb3a6f6-4545-41b9-bcbd-6b76f079f28b · outbound

This paper cites Neural fitted Q iteration–first experiences with a data efficient neural reinforcement learning method.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Neural fitted Q iteration–first experiences with a data efficient neural reinforcement learning method

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.490087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.835771Z digest=sha256:e9e902d6687f4abf6c5d19546793222ccc2d02fbacc8f6487245952e53af469b

Observation e8529386-077f-48d3-aff5-270c99a98bdd · outbound

This paper cites Deep exploration via bootstrapped dqn.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Deep exploration via bootstrapped dqn

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.839747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.839747Z digest=sha256:1d13201347f5d19655dd8bd88896f04e1bdc5b1989dd97cd84387f0291206807

Observation 437aba3e-ae67-4b77-9ca5-84988a9186ab · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Playing Atari with Deep Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.843382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.843382Z digest=sha256:f94bd9aa7c6a4d3f788e9ac3123aa130a6120f586a8610308d9aec5b6069ff72

Observation 93088779-e9d7-477c-9357-8f8d585e4d46 · outbound

This paper cites Q-learning with linear function approximation.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Q-learning with linear function approximation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.469419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.847237Z digest=sha256:9401af85a55ca1fb8c0d32e4a3b2e4410953920879d3f09e241c113d14230c6a

Observation c4b897ad-5a0e-468c-a703-3d727afdbec3 · outbound

This paper cites Sample complexity of reinforcement learning using linearly combined model ensembles.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Sample complexity of reinforcement learning using linearly combined model ensembles

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.438526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.854287Z digest=sha256:2a8f02a5fabea2d770813e1c3d978d617b621d795cb21b8923e5d3e4a14337e2

Observation 2cdac10e-b52c-4cb3-aeb4-900e4b5f0a31 · outbound

This paper cites Model-Ensemble Trust-Region Policy Optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Model-Ensemble Trust-Region Policy Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.858060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.858060Z digest=sha256:017f51b7bfcca9aa6d66c7b67f7b6efd9166f940fcec61cecc688f43d899710e

Observation 9fc40cfa-629d-46e3-adf7-3a93dd311a94 · outbound

This paper cites Deep reinforcement learning in a handful of trials using prob- abilistic dynamics models.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Deep reinforcement learning in a handful of trials using prob- abilistic dynamics models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.862136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.862136Z digest=sha256:d9dd8eabf843c7f4346d3e710576ae425c1f9dc5bc9dc32d372af3e2e379ca76

Observation bf293276-88c3-4155-ad35-927b6832fecf · outbound

This paper cites Asynchronous methods for deep reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Asynchronous methods for deep reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.865881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.865881Z digest=sha256:de59dba42beaf4da08d7d90f2e7a46f7a13cf9dd6dc470dda9942fb2de502d91

Observation 35e61a55-621d-4e71-9e1b-795e9e89febf · outbound

This paper cites Ensemble link learning for large state space multiple access communications.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Ensemble link learning for large state space multiple access communications

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.392417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.870066Z digest=sha256:193e341084e8fca795de37580a07c684b8a5dc07adc7151bc3ded61dad5b0b61

Observation f911ecda-c405-4234-bd60-bd8ea2784f04 · outbound

This paper cites Ensemble graph Q-learning for large scale networks.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Ensemble graph Q-learning for large scale networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.377898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.873712Z digest=sha256:eeb4a9ac1558138954381c32bf615429b5a54258122209ae83baee870e12ecda

Observation f1f5829f-c46c-415a-9af1-6c27bc24101b · outbound

This paper cites Multi-timescale ensemble q-learning for markov decision process policy optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Multi-timescale ensemble q-learning for markov decision process policy optimization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.363970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.877389Z digest=sha256:537f8f7e3e5637e0482b2a3ab30acc118d39429eddd6a75646c9416a4d126ee5

Observation c4653aeb-3949-4bef-acfc-97fae01d8eb2 · outbound

This paper cites Leveraging digital cousins for ensemble q-learning in large-scale wireless networks.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Leveraging digital cousins for ensemble q-learning in large-scale wireless networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.348408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.882138Z digest=sha256:b69b2b3c44ae6c191917a30681397cbb87a227b5d66687c6d51ccbe1509a8e14

Observation f7c8391b-7c76-480b-a1b7-38c697c05aa4 · outbound

This paper cites A novel ensemble q-learning algorithm for policy optimization in large-scale networks.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning A novel ensemble q-learning algorithm for policy optimization in large-scale networks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.335126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.885810Z digest=sha256:f868ebec5953bda9f8b43d0e9c19cbecd64dc5d36aa607aeb91023b453722623

Observation 94fe94b0-4fbe-4b16-9f65-7eb442b8d7c1 · outbound

This paper cites Link analysis for solving multiple- access mdps with large state spaces.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Link analysis for solving multiple- access mdps with large state spaces

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.322383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.889556Z digest=sha256:cd02a2e34fbd6f2dadcaeb36091bd584a851d34b74386fe02327f04a4f115a97

Observation d3692961-f7ac-4542-aa87-34bbb7cef735 · outbound

This paper cites Is q-learning provably efficient? Advances in neural information processing systems, 31, 2018.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Is q-learning provably efficient? Advances in neural information processing systems, 31, 2018

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.893407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.893407Z digest=sha256:382a4117ba0b5c30d81e0fa7edd16951b4053d9716a793dfb64b835aac5bc8ab

Observation 8d6795f2-3ec8-4547-b960-5b58ae3ea6ff · outbound

This paper cites Online robust reinforcement learning with model uncertainty.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Online robust reinforcement learning with model uncertainty

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.897232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.897232Z digest=sha256:b234f3f4f6f88e9f8cf4fa5a1f5d923dc0b51bf2ac80eced55789ebfcb333c82

Observation da4cd2ce-b9f7-437a-a4a1-ddeb4688a324 · outbound

This paper cites Conser- vative q-learning for offline reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Conser- vative q-learning for offline reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.285828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.901220Z digest=sha256:dec432a0a6a5d3fbfa7e34b1a4a2ad3065acb4401102638fab8616349d8a6fe6

Observation f0ea9ff9-5f10-44ff-b982-5d7ca4bfe099 · outbound

This paper cites An optimistic perspective on offline reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning An optimistic perspective on offline reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.273193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.905076Z digest=sha256:f586088d4120d659405d85bba557d5f6a2d7ee1ecb1e1dc2638237c42fb4314f

Observation 2b6f2766-2cab-49cd-adc5-e74516f10fdb · outbound

This paper cites Leveraging offline data in online reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Leveraging offline data in online reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.260691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.908840Z digest=sha256:32c53bbd410d1534a6ac165c33a3bc2aad8fd07c02335acfab9654378064283a

Observation 1b4cfa2c-1a0c-4510-bd0b-10d85210e379 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.912632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.912632Z digest=sha256:8548f9eee09ad228477654ff8690a99803e0d9ff4d60e58e8d0121f0b59a630e

Observation a7cf1705-e583-4563-811c-3aa513f289a5 · outbound

This paper cites Policy finetuning: Bridging sample-efficient offline and online reinforce- ment learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Policy finetuning: Bridging sample-efficient offline and online reinforce- ment learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.247073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.916639Z digest=sha256:9687557116abd7b1ddbcdf960d85282d8cb56d5e4a56e665b68e0030a06d3bc5

Observation 2e6d1209-d8cc-4956-aa2c-08bc1c594bdc · outbound

This paper cites Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.920591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.920591Z digest=sha256:eb3d43fd205d30c1079ea1c20061e2734b03e34bc0cb7bdd1959c11bf3500e82

Observation af046d29-91fc-40f3-9395-0c54ed904789 · outbound

This paper cites Com- paring exploration strategies for q-learning in random stochastic mazes.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Com- paring exploration strategies for q-learning in random stochastic mazes

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.230930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.924650Z digest=sha256:708f8ba22233b60ce57fb0ca25d7f258abc1d2cd7f8379c909978c45c4bfc403

Observation b10298f2-4e93-40ef-864d-5fcf11b9e2d3 · outbound

This paper cites The Role of Coverage in Online Reinforcement Learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning The Role of Coverage in Online Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.928379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.928379Z digest=sha256:20b44be7cc9fec74664320e143e42b34060d80ea537790dd38a639e91d507342

Observation 17fc7772-cfc9-4e89-83f4-c044a97f4ba1 · outbound

This paper cites Pessimistic Model-based Offline Reinforcement Learning under Partial Coverage.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Pessimistic Model-based Offline Reinforcement Learning under Partial Coverage

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.932746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.932746Z digest=sha256:f73635697fced123a313b79b07331d385eb2cde85b198fcac6f6167b77eb3b14

Observation 0994f7b3-a7db-4031-a407-46709bcebc5f · outbound

This paper cites What can online reinforcement learning with function approximation benefit from general coverage conditions? In International Conference on Machine Learning, pages 22063–22091.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning What can online reinforcement learning with function approximation benefit from general coverage conditions? In International Conference on Machine Learning, pages 22063–22091

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.211572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.936660Z digest=sha256:41533beda7e13ec55435fb3d7b8c46b8093774719c57c8a8a705c37852e7b3c1

Observation 14212655-35fd-4a34-ac2e-4f20997ff5ef · outbound

This paper cites When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.196846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.940435Z digest=sha256:cd706f0b0fb12c1553691103b46c7549c86c35623262365f1a2b025c899c460d

Observation 4a951e64-b95d-4ded-a363-bd8dd1f1a35d · outbound

This paper cites Coverage analysis of multi- environment q-learning algorithms for wireless network optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Coverage analysis of multi- environment q-learning algorithms for wireless network optimization

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.174530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.944097Z digest=sha256:c45c32c4d8beeff2bfb016da21c610dd80a23d5d0d5a5d0d6d6e73aa0a9645c7

Observation 62ba41a5-6357-4d1d-b602-938ee6d2ab46 · outbound

This paper cites Convergence of Q-learning: A simple proof.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Convergence of Q-learning: A simple proof

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.160552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.947548Z digest=sha256:f89c5de32d4d38d91b4a7973f72d63fa394f718efb0b726edc7cc29ee5e80cb1

Observation 1afabcb6-37a8-4099-9fc0-3406415ba7bf · outbound

This paper cites Issues in using function ap- proximation for reinforcement learning.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Issues in using function ap- proximation for reinforcement learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.140862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.951314Z digest=sha256:7a71d032e09e53729673e2b6577d1af6833ee3b30b68a7787970b2c91fdc03b1

Observation 244acb39-d088-4af2-aa02-2e23dee15268 · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.955046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.955046Z digest=sha256:b5091ac941c1dde4fb6de85359186ebc5e8267799d9ab79394541fe20e54d668

Observation eadff9d3-985d-40c0-8e74-1cf2992aabf5 · outbound

This paper cites The virtues of laziness in model-based rl: A unified objective and algorithms.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning The virtues of laziness in model-based rl: A unified objective and algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.959107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.959107Z digest=sha256:d20d87a5b77135e74c3ab03a36f86fbd09f98ecefe18c65a9393d16fd542c48f

Observation e72cc9b0-35ce-41e7-9d18-6532afba8694 · outbound

This paper cites A Multi-Agent Multi-Environment Mixed Q-Learning for Partially Decentralized Wireless Network Optimization.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning A Multi-Agent Multi-Environment Mixed Q-Learning for Partially Decentralized Wireless Network Optimization

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T21:48:17.003329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.963053Z digest=sha256:57c80452b012c1cf625e1b89d88aa562353d6ba97a36af81554a1c0c3f8f7b73

Observation 4dcbefa6-4607-4d8d-8b7f-c1537b30145f · outbound

This paper cites Springer, 2007.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning Springer, 2007

Reference 2007

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:48:17.455306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T21:48:16.850593Z digest=sha256:15241b6b309ba645826844449630fe81d49d96011aaea7c6abb9b89e0dbefe29

Pith citing papers

No inbound Pith citation observations are available.