Pith. sign in

Paper Citation Record · LEDGER

Near-Optimal Sample Complexity for MDPs via Anchoring

As of 10 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 2 inbound Pith citation observations for arXiv:2502.04477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04477 v2

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:48:33.114297Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T06:51:04.076537Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact1
  • verified fuzzy47
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6c0d75f-e4d1-4cff-afac-9ec6db3d2ee0 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:35.381305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.777072Z digest=sha256:975846333d3b92529d9009602d99e0af833a6ea9057bf71faa518e158630c8d1

Observation f580cb8d-3ce7-47d6-8e4d-82c1326750de · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:35.366170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.801994Z digest=sha256:b349733213a2af897f5d71c5d5d27265d57f0baa5a76a4f93abcb1964f8bc08f

Observation 34ab2e25-ed2b-4733-b1cd-dac981e81cda · outbound

This paper cites G., Munos, R., and Kappen, H.

Near-Optimal Sample Complexity for MDPs via Anchoring G., Munos, R., and Kappen, H

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.351423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.825800Z digest=sha256:96eb497c6c28a9ed738725d81ac82ac4908ad867a6a12e6d088d1c0ee2c7e0c5

Observation 8132e9a9-7de1-479e-b587-902eb5ede860 · outbound

This paper cites Weighted sums of certain dependent random variables.

Near-Optimal Sample Complexity for MDPs via Anchoring Weighted sums of certain dependent random variables

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.337490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.844624Z digest=sha256:a26190341c7599bb8052f9fc9587eea40ff7e463ec3fcfb5c50fdd552ac12607

Observation f4227128-9771-420b-a431-f092f2d1e7b8 · outbound

This paper cites U., and Aggarwal, V.

Near-Optimal Sample Complexity for MDPs via Anchoring U., and Aggarwal, V

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.322700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.851877Z digest=sha256:ae722b04b6dc65c24a274406c6c66b24fa60b390d1ca81bc044a8cbe234fd12a

Observation 5471941a-dabe-4723-9730-56588789be6e · outbound

This paper cites A M arkovian decision process.

Near-Optimal Sample Complexity for MDPs via Anchoring A M arkovian decision process

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.308459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.856167Z digest=sha256:e788af176852f94da509797152066395c1a59c27313ca35f866de49d00c5729e

Observation 8d9953e6-9138-49bf-bd63-e02467ea3586 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:35.294417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.861333Z digest=sha256:5502ad3a60393961084e90631d9f9445530fc03ce0d5e862376ef949b00a9fbe

Observation b1c8853a-bd7e-45d4-8c21-f349c5a20cf7 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:35.281233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.866004Z digest=sha256:02cf8cdbbe92a87949f054b520a80f723b6e258df9e55a2b5a91b62b984e04d4

Observation 079d0649-fff0-47fb-924c-91ab2a79dfa8 · outbound

This paper cites What Doubling Tricks Can and Can't Do for Multi-Armed Bandits.

Near-Optimal Sample Complexity for MDPs via Anchoring What Doubling Tricks Can and Can't Do for Multi-Armed Bandits

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T22:48:31.870679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:48:31.870679Z digest=sha256:70dbf1f2f69f23739de9971f275a42366cc34c22a6c58e2634e55b198439f2a7

Observation ab8a314a-4571-4704-ad5e-8472e4c42875 · outbound

This paper cites Discrete dynamic programming.

Near-Optimal Sample Complexity for MDPs via Anchoring Discrete dynamic programming

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.265306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.875465Z digest=sha256:bfa845d0f56bdf34e0365d6d5028f9624ef4a5afaa5d5248f46b92fba4367c45

Observation bc3369f8-5ca6-47e8-8396-7388b88bff3f · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:35.141069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.879796Z digest=sha256:e869e3250bbd0fdda20278c3c30496226eb0c8c356da2fc064897be6ca0dd8d1

Observation 4e1ec5f7-cd78-4153-b063-010211071eb6 · outbound

This paper cites Stochastic Halpern iteration in normed spaces and applications to reinforcement learning.

Near-Optimal Sample Complexity for MDPs via Anchoring Stochastic Halpern iteration in normed spaces and applications to reinforcement learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-08T22:48:33.221284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.884881Z digest=sha256:d9f5cca90b3d93e0b691e3168ea0b9535070a12326db6015fdfb91caa2d35355

Observation 7d84cbcd-7415-4512-af74-d13c7776f005 · outbound

This paper cites Stochastic H alpern iteration with variance reduction for stochastic monotone inclusions.

Near-Optimal Sample Complexity for MDPs via Anchoring Stochastic H alpern iteration with variance reduction for stochastic monotone inclusions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.043555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.888965Z digest=sha256:71dc4ffe5a6627bf3d45cca23b42b2e46600eb312dd47139143dbbd7923209af

Observation 994e4a04-f101-452c-bcb0-7d0cc07c9f33 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:35.028925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.893704Z digest=sha256:f35098cd0e9764a8735998920b314332fba0043467e83041a008251e8b4bd299

Observation 1a5c1a4f-ee6b-4a6c-b13e-786381d12c5d · outbound

This paper cites Average-reward model-free reinforcement learning: a systematic review and literature mapping.

Near-Optimal Sample Complexity for MDPs via Anchoring Average-reward model-free reinforcement learning: a systematic review and literature mapping

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T22:48:31.897502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:48:31.897502Z digest=sha256:1bd79a23f1b65bb1adef655ccbcb808b4ab7aa3f56dfef986c2782db460aa388

Observation 8a31b2dc-d673-4cbe-b1fb-54666603e240 · outbound

This paper cites Tree-based batch mode reinforcement learning.

Near-Optimal Sample Complexity for MDPs via Anchoring Tree-based batch mode reinforcement learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.015122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.902132Z digest=sha256:1f606afb6e509bbd22cab76d6be9a3c42428d88925c0c5ed6b8fbff36457c679

Observation 3bc3b9aa-0e27-406f-b71c-8920beb186ff · outbound

This paper cites Regret minimization in MDP s with options without prior knowledge.

Near-Optimal Sample Complexity for MDPs via Anchoring Regret minimization in MDP s with options without prior knowledge

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:35.000334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.906920Z digest=sha256:b3a037efbc74aaf90aa103690bcb764c817731c211d6eb32c49073550038c887

Observation 43b847e9-ef88-4276-9a79-78a4d5848601 · outbound

This paper cites Fixed points of nonexpanding maps.

Near-Optimal Sample Complexity for MDPs via Anchoring Fixed points of nonexpanding maps

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.985584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.910727Z digest=sha256:e63596c28232e15e146adedb97fd284583e714b2ae1c11dd0d83cd0e7f1e585b

Observation 8660ebd0-bde8-4f66-8347-a587bb714912 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:34.972046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.915182Z digest=sha256:63928e847d3a3a2fc034a39e7f445fa10fc83b6bec99d1466d5f8d01f0c77673

Observation c15d6450-9dd2-4201-99a7-09a588487e6c · outbound

This paper cites and Sidford, A.

Near-Optimal Sample Complexity for MDPs via Anchoring and Sidford, A

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.958140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.920096Z digest=sha256:b80ae7719e235fdd862d41dd02e4a7204ee889cff4156961f7fbdd5b99b592ad

Observation 4ee15322-a920-4172-97af-f85c87a51f22 · outbound

This paper cites and Sidford, A.

Near-Optimal Sample Complexity for MDPs via Anchoring and Sidford, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.944742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.924979Z digest=sha256:4a49844e79124bb256992a43f204f15155b4de67e7c940b8158bb6bef9f5103b

Observation 81b6d8ee-d219-4423-9eba-73ba8387d9bc · outbound

This paper cites Feasible Q -learning for average reward reinforcement learning.

Near-Optimal Sample Complexity for MDPs via Anchoring Feasible Q -learning for average reward reinforcement learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.931061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:31.963818Z digest=sha256:1636111e4cc79240fad8f9f8a48c00816140529c6a03a250cf6c8995ef34abfe

Observation 0a4f39db-4299-4eea-b331-fb249c65830a · outbound

This paper cites Truncated variance reduced value iteration.

Near-Optimal Sample Complexity for MDPs via Anchoring Truncated variance reduced value iteration

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.917313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.065918Z digest=sha256:a6f015be787efe96e62f8328f86590b227b6ef8a58d38aaf67bf8e4ec15f7b3f

Observation 7e47a3ba-f557-421c-acfe-ccc5761b3685 · outbound

This paper cites and Zhang, T.

Near-Optimal Sample Complexity for MDPs via Anchoring and Zhang, T

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.798170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.177036Z digest=sha256:fc59071437aec2f49093c515f75d9b2df101cb87ce936ca5fe982e73f5dfaeb9

Observation aefbacb3-3aee-4b3b-a896-8069f2aa383a · outbound

This paper cites An -best-arm identification algorithm for fixed-confidence and beyond.

Near-Optimal Sample Complexity for MDPs via Anchoring An -best-arm identification algorithm for fixed-confidence and beyond

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.653622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.270625Z digest=sha256:39ded3fa135776b052a021fab1edd916ae960bb64715c76c5ca50c051d509032

Observation 97798eca-1f4e-44f7-938f-a5c9d30cf669 · outbound

This paper cites and Jamieson, K.

Near-Optimal Sample Complexity for MDPs via Anchoring and Jamieson, K

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.632874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.309190Z digest=sha256:52c7c19f260e63ed337d3fa60d650bcb7467289525491070843a1bfd84d95291

Observation 9c44ef31-2546-4bb7-9763-ba3ad9274521 · outbound

This paper cites and Singh, S.

Near-Optimal Sample Complexity for MDPs via Anchoring and Singh, S

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.619684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.356479Z digest=sha256:ec5885146d47160fc47d48432d40b9be21778fb8f36d8d75430232b59e0c78f8

Observation 5966d6d3-b39a-47f9-84a2-7b36710a261d · outbound

This paper cites Accelerated proximal point method for maximally monotone operators.

Near-Optimal Sample Complexity for MDPs via Anchoring Accelerated proximal point method for maximally monotone operators

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.604718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.392996Z digest=sha256:89f939c9e3042fa0d09c150d94db06cdbc960d51fe78c57ed01b5d51fb2f48af

Observation 8e19db17-a48f-4f7d-8968-1d960b533fb2 · outbound

This paper cites Y., and Mannor, S.

Near-Optimal Sample Complexity for MDPs via Anchoring Y., and Mannor, S

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.590069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.415711Z digest=sha256:2676298f6b105072284d7e389d3d8e025ba13d8c65a1148b5a08203cc6bc4bb4

Observation ae9d1a2d-246b-4274-83e5-dce8ac801723 · outbound

This paper cites Y., Srikant, R., and Mannor, S.

Near-Optimal Sample Complexity for MDPs via Anchoring Y., Srikant, R., and Mannor, S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.576590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.427252Z digest=sha256:3d9a655d5259fcb0accac983c3d36c0eaf55f0c20b8f72951ca484389376d071

Observation a5f05752-a821-416e-a600-24d01b6ff205 · outbound

This paper cites and Ryu, E.

Near-Optimal Sample Complexity for MDPs via Anchoring and Ryu, E

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.562153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.431951Z digest=sha256:2cb31170e4529f9444502971589dde028dedf73cd616d0d389d5830349e8c460

Observation 68678cd1-5010-4abf-9825-204bdcf66f00 · outbound

This paper cites and Ryu, E.

Near-Optimal Sample Complexity for MDPs via Anchoring and Ryu, E

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.547674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.437826Z digest=sha256:4c6fe3eeb511b3492a688d00500bd5ebeed3399fb322246ccf9ee0488c3258c6

Observation 5ea594d2-ff32-4602-b069-f2c62a3b3ad2 · outbound

This paper cites Breaking the sample size barrier in model-based reinforcement learning with a generative model.

Near-Optimal Sample Complexity for MDPs via Anchoring Breaking the sample size barrier in model-based reinforcement learning with a generative model

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.506560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.442294Z digest=sha256:367f9d6eaf34e30148fb3198e2c480e02741527361a897ee7527b5a384182ad9

Observation 246097ab-b289-4951-86d7-c878ef346b83 · outbound

This paper cites Stochastic first-order methods for average-reward M arkov decision processes.

Near-Optimal Sample Complexity for MDPs via Anchoring Stochastic first-order methods for average-reward M arkov decision processes

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.384186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.446716Z digest=sha256:808458b7e0aee8a826ffd9e6f605516277ca17639a4bbdf626b8feb16902decc

Observation 67877b08-3ef0-4f46-a86b-7cf26beb410a · outbound

This paper cites On the convergence rate of the H alpern-iteration.

Near-Optimal Sample Complexity for MDPs via Anchoring On the convergence rate of the H alpern-iteration

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.339419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.451136Z digest=sha256:24894d74703a616762b2441bbf27065bd4dd289c5c3a418f5db5ed64b21e37ea

Observation 9ee31ff6-ffcf-4c41-999b-b86b80114a7c · outbound

This paper cites Average reward reinforcement learning: Foundations, algorithms, and empirical results.

Near-Optimal Sample Complexity for MDPs via Anchoring Average reward reinforcement learning: Foundations, algorithms, and empirical results

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.325611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.455356Z digest=sha256:9d89915d19663254e2a1bc3f87080336d019f580e535cc263586f7a5f70ec12c

Observation 7d047367-15e1-4d05-a53d-154aaae69599 · outbound

This paper cites Finding all -good arms in stochastic bandits.

Near-Optimal Sample Complexity for MDPs via Anchoring Finding all -good arms in stochastic bandits

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.311312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.459849Z digest=sha256:0a254633891fee2ceeaa79cedb676bc44c6517c15fbb3da160605c153fdb5ace

Observation b1a61167-9a7b-47a2-811a-24713a39dd31 · outbound

This paper cites A., and et al.

Near-Optimal Sample Complexity for MDPs via Anchoring A., and et al

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.296661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.463764Z digest=sha256:e3adabca5153946eb71164b0870f06f39739bc75f5ae16afe0a02d1a5c4df0e5

Observation 3ced112a-25b0-4238-967a-97e59a574c06 · outbound

This paper cites and Szepesv \'a ri, C.

Near-Optimal Sample Complexity for MDPs via Anchoring and Szepesv \'a ri, C

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.284077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.468453Z digest=sha256:b170ce2b22a64a82c180059d9848af468385e6289deb5982aa26a575d0d1e615

Observation ee982dfd-3408-430c-b656-ae1a0c6d73a5 · outbound

This paper cites and Srikant, R.

Near-Optimal Sample Complexity for MDPs via Anchoring and Srikant, R

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.270618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.472967Z digest=sha256:cf9496a65ff439092a44965b7f643ff4787ad26f8393ee4013ef4f1bbf672aac

Observation f30b2363-3cca-4118-8b1e-2d82de96dd01 · outbound

This paper cites and Okolo, N.

Near-Optimal Sample Complexity for MDPs via Anchoring and Okolo, N

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.251721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.477471Z digest=sha256:f5fc62a11214913f6b2646c3a83fb4e13e04217dbbbed8c23d80441141b747ea

Observation c79f08b4-a558-4c76-ab8f-b6a0451e358b · outbound

This paper cites M., Liu, J., Scheinberg, K., and Tak \'a c , M.

Near-Optimal Sample Complexity for MDPs via Anchoring M., Liu, J., Scheinberg, K., and Tak \'a c , M

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.236817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.494634Z digest=sha256:f456c5d5e24f0ee3536ac1a6d5d736f8a4e7efaf21bb03cda68b794c7f21981f

Observation 7f37755b-22d8-443e-9284-ade197858d02 · outbound

This paper cites and Ryu, E.

Near-Optimal Sample Complexity for MDPs via Anchoring and Ryu, E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:34.208132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.588998Z digest=sha256:4457b07134e3461a3370f564d4c7d6ae84cb345745871ca19f8c1ef2763df158

Observation df731c54-283e-4cb8-a503-e51459b80af9 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:34.108545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.734829Z digest=sha256:c092dd6ad6816e139b908b4f8ac526cf1a2b29bacae5e5a127622947a4705354

Observation 6c8ed0f4-ec9e-40c0-a5b5-40130bdec4b8 · outbound

This paper cites Strong convergence theorems for resolvents of accretive operators in B anach spaces.

Near-Optimal Sample Complexity for MDPs via Anchoring Strong convergence theorems for resolvents of accretive operators in B anach spaces

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.999717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.748479Z digest=sha256:140faac6db31eff84c9456f60694afe3c0a75df6adc2ac5d8ac4c319ac9ebdab

Observation 7fab3d07-0926-4b1f-8507-88f7db287c0f · outbound

This paper cites and Mansour, Y.

Near-Optimal Sample Complexity for MDPs via Anchoring and Mansour, Y

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.934187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.768186Z digest=sha256:8ebe34174711706145e59f1a8c4b06c0f67b6c3a427c99ddd1bad497b2a8a929

Observation 312fa99d-b052-4dd4-a9ad-f6e3ac60ac59 · outbound

This paper cites and Shtern, S.

Near-Optimal Sample Complexity for MDPs via Anchoring and Shtern, S

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.920700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.780734Z digest=sha256:e4fedb283556aec02283fa5d9d92317e75aa5166eda8f0109a0a03237205f4fd

Observation 1c0aeceb-2f32-4a56-8a46-c5f6ab9b45b8 · outbound

This paper cites Near-optimal time and sample complexities for solving M arkov decision processes with a generative model.

Near-Optimal Sample Complexity for MDPs via Anchoring Near-optimal time and sample complexities for solving M arkov decision processes with a generative model

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.906432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.788355Z digest=sha256:7b81eb0a61e4705812fa67924f9ec6fde4b9a05ce7989a925a16b1acc55b172f

Observation c4dbc32f-8889-4c84-b184-281d760b7b39 · outbound

This paper cites Variance reduced value iteration and faster algorithms for solving M arkov decision processes.

Near-Optimal Sample Complexity for MDPs via Anchoring Variance reduced value iteration and faster algorithms for solving M arkov decision processes

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.891913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.793396Z digest=sha256:502aa5d6ae3c42539f163921d5a7ed5bece06f3efdb2d093139490e1546363d8

Observation 1148d948-2309-453f-a4fd-c04a3f99b1d1 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:33.877940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.797742Z digest=sha256:1be8c5fce07fef79e0e6087c979cc0a2c615c86f751dac49005f64d97e71dc2c

Observation bf4e78f9-d24d-4f6a-a270-022d544ba479 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:33.839562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.801806Z digest=sha256:221e701a366f6d5d8ef39de0e8cb3598088a285ef0df1ffdedaa8250aa08f578

Observation 9cf865e6-99bd-4526-8fb5-490dc7fcc371 · outbound

This paper cites Algorithms for Reinforcement Learning.

Near-Optimal Sample Complexity for MDPs via Anchoring Algorithms for Reinforcement Learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.701272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.806678Z digest=sha256:2427d45f7cae952ebfe7dbec19fc7a6e150bf147238b7ed17588223ded804c21

Observation c79fe039-f5f9-4024-8dc4-f355a07aea7b · outbound

This paper cites Finding good policies in average-reward M arkov decision processes without prior knowledge.

Near-Optimal Sample Complexity for MDPs via Anchoring Finding good policies in average-reward M arkov decision processes without prior knowledge

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.591673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.810601Z digest=sha256:bf061be8afc53956280c7af1a82bccd7f2894e62f9228a087d5a3066e5745bc6

Observation a7e16fd6-1b54-4c91-a3da-ba2c5e5aec17 · outbound

This paper cites Variance-reduced $Q$-learning is minimax optimal.

Near-Optimal Sample Complexity for MDPs via Anchoring Variance-reduced $Q$-learning is minimax optimal

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T22:48:32.814848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:48:32.814848Z digest=sha256:73b74db99ab4a031505afd2168153f9191d2c643855ab59ee551337cd225761c

Observation 807d7a90-8441-4ab6-af08-1f3a2e150258 · outbound

This paper cites an unresolved cited work.

Near-Optimal Sample Complexity for MDPs via Anchoring Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:48:33.576996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.819248Z digest=sha256:d3e0310dca32c3b03b4147273f2370f703ab02d0596901434f996f31d38b4bcf

Observation 32ab1a26-20b5-443f-9542-807e6181f9e3 · outbound

This paper cites Near Sample-Optimal Reduction-based Policy Learning for Average Reward MDP.

Near-Optimal Sample Complexity for MDPs via Anchoring Near Sample-Optimal Reduction-based Policy Learning for Average Reward MDP

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T22:48:32.824182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:48:32.824182Z digest=sha256:01baba54260fbdecff37743c34e3481b48c43cc23cfddba90410a48e53ec2ae1

Observation c5337e60-d0f6-445d-9e53-ee7c5761380f · outbound

This paper cites Primal-Dual $\pi$ Learning: Sample Complexity and Sublinear Run Time for Ergodic Markov Decision Problems.

Near-Optimal Sample Complexity for MDPs via Anchoring Primal-Dual $\pi$ Learning: Sample Complexity and Sublinear Run Time for Ergodic Markov Decision Problems

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T22:48:32.829460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:48:32.829460Z digest=sha256:165cfcb7beafec5bedd681310c4f0cbb2d91547a227330692dce2c95db5bd0ba

Observation 086120f1-6b5a-43c6-8f19-87d59acf8650 · outbound

This paper cites Optimal sample complexity for average reward M arkov decision processes.

Near-Optimal Sample Complexity for MDPs via Anchoring Optimal sample complexity for average reward M arkov decision processes

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.562637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.834124Z digest=sha256:578c9596cf8456b8c7e962537a882671b1bf6d1dc67efc24de9fc303420ee96e

Observation c3b98127-31c4-4daf-a83b-0494e20a430b · outbound

This paper cites J., Luo, H., Sharma, H., and Jain, R.

Near-Optimal Sample Complexity for MDPs via Anchoring J., Luo, H., Sharma, H., and Jain, R

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.547358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.838129Z digest=sha256:5d1e534a0e83ec8bef80fedb2cdb63484235edd55fb3e912c263d0c5c9e8a6fd

Observation 3f56888e-d312-472d-b92b-47e14efd51e5 · outbound

This paper cites Approximation of fixed points of nonexpansive mappings.

Near-Optimal Sample Complexity for MDPs via Anchoring Approximation of fixed points of nonexpansive mappings

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.519795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.842808Z digest=sha256:8b31d1dcec7c499f6d65211cbe1f9d166fd6cf4342825dddca7547688da9a42d

Observation f13a70e5-2694-4ef6-a9cd-5043d6e698a2 · outbound

This paper cites Iterative algorithms for nonlinear operators.

Near-Optimal Sample Complexity for MDPs via Anchoring Iterative algorithms for nonlinear operators

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.353677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.847597Z digest=sha256:f3c80da4e1de84e9b900db2db40959449966ed9add7159cfd767e486a4cdb9d8

Observation 0f84857f-5ec0-46da-ab79-1cd7eb8ea5a7 · outbound

This paper cites and Ryu, E.

Near-Optimal Sample Complexity for MDPs via Anchoring and Ryu, E

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.294665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:32.884091Z digest=sha256:72227132be2238c12c9eefcad73de95f97f8cceb1056f1eaa06ef8c993dcd79c

Observation 0a693ff1-51e2-4bad-a79c-5b7160dde0a8 · outbound

This paper cites and Xie, Q.

Near-Optimal Sample Complexity for MDPs via Anchoring and Xie, Q

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.279188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:33.022441Z digest=sha256:fb566424c15aebb0a4d87ee9451e473f1333c86a360708c9e07847f02233c657

Observation 75878761-0f78-4812-83da-89c6fd3bee47 · outbound

This paper cites and Chen, Y.

Near-Optimal Sample Complexity for MDPs via Anchoring and Chen, Y

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:48:33.263816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T22:48:33.109820Z digest=sha256:196039fe6203b2c3a7e7874be51e15d66936becd3d65ed48ca30b18e9a8df91f

Observation 0fe317f3-ab06-4ec0-9f6a-226958058625 · outbound

This paper cites write newline.

Near-Optimal Sample Complexity for MDPs via Anchoring write newline

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T22:48:33.114297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:48:33.114297Z digest=sha256:382baf07e35f68f6845ed64f838cb4399bbc2d0c73c35905e5b382db85e1ca6b

Pith citing papers

Observation 57ea9f20-2673-44da-aa4a-4938303e4ca3 · inbound

Auto-exploration for online reinforcement learning cites this paper.

Auto-exploration for online reinforcement learning Near-Optimal Sample Complexity for MDPs via Anchoring

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T18:18:58.439629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:18:58.439629Z digest=sha256:2f9ece665c964dc57ea5f19c9c9286478211650f4c5b25d482c05deb505dfa33

Observation fd92dba0-5cbd-4100-8480-4818740c509a · inbound

Auto-exploration for online reinforcement learning cites this paper.

Auto-exploration for online reinforcement learning Near-Optimal Sample Complexity for MDPs via Anchoring

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T06:51:04.076537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:51:04.076537Z digest=sha256:6270de9b23f15e5b7ca470aef01a5ef5b4126f1dd73c2cf1a9896c4021dfd83e