Pith. sign in

Paper Citation Record · LEDGER

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions

As of 16 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 0 inbound Pith citation observations for arXiv:2608.06545.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06545 v1

Coverage vector

measured 100 of 300 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:39:15.637632Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 300 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved94
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2ccdd33f-3f04-4beb-a6ae-793c30b353ad · outbound

This paper cites Towards Tight Bounds on the Sample Complexity of Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Towards Tight Bounds on the Sample Complexity of Average-Reward

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.203436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.203436Z digest=sha256:f58589fd7b414718af4c4d972f865ab519d1fb21d464262d84a8a3c78dbe3421

Observation f2cc61b9-093d-4e5a-a0a2-8c15624c0184 · outbound

This paper cites Foundations and Trends in Machine Learning , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Foundations and Trends in Machine Learning , volume =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.207525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.207525Z digest=sha256:c13135d9eb0ab5cfb9117337173a1fd9693373d55fd999a32e0569abce60e37f

Observation 8ea6c8a4-068c-48d2-9b23-c30a2867e586 · outbound

This paper cites Operations Research , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Operations Research , year =

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-15T14:39:17.218618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.211143Z digest=sha256:57836dd8d481a024a3aa72ae3c96977a37f81f2aa21c19a26402db173f3e1029

Observation 18cd24de-c12a-4c63-8b12-ff2875b2223e · outbound

This paper cites Near Sample-Optimal Reduction-Based Policy Learning for Average Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near Sample-Optimal Reduction-Based Policy Learning for Average Reward

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.215035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.215035Z digest=sha256:9f97ff3fe82249fd74203cdd48e5afb397233ca5d732ec6735cad2060be529d4

Observation 308a30ac-f493-43a3-bd7b-cc343280fed5 · outbound

This paper cites Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.218684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.218684Z digest=sha256:e56edc612c75a1bcf847ba6f80edfecfb8ad5cceaf5caf206e8cd0338930c174

Observation 07ceabbc-b080-41b6-adf4-74793e9a4c3b · outbound

This paper cites and Tewari, Ambuj , booktitle =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions and Tewari, Ambuj , booktitle =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.222526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.222526Z digest=sha256:e1cd54a76f297b086230c9996d8c1a8b99661c2e7de6d7a535449f5071722098

Observation bf3c2cea-2013-4688-bb5a-c88449183c95 · outbound

This paper cites Proceedings of the 35th International Conference on Machine Learning , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Proceedings of the 35th International Conference on Machine Learning , pages =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.225630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.225630Z digest=sha256:49f7a9f16181f59b6b24c5f002252cfb721ed3fbac364ee6c0fead6d83cb9ceb

Observation a9581ca9-bca9-44e1-b18c-66313c1fc4a2 · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Proceedings of the 42nd International Conference on Machine Learning , pages =

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.229174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.229174Z digest=sha256:c237de6f212da25f8d124eaa5e81096be0363c96a67d5504130c45592d95456b

Observation d541f5cd-ef54-46b1-8f06-ffb1343a118f · outbound

This paper cites Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.232280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.232280Z digest=sha256:0cecf9813e2a3fae49e09093332f558e393ba49ae2f827908a674fa94471f454

Observation e4a6617e-4655-40c1-86f7-3fdbac929a7e · outbound

This paper cites High-Dimensional Statistics: A Non-Asymptotic Viewpoint , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions High-Dimensional Statistics: A Non-Asymptotic Viewpoint , year =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.235726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.235726Z digest=sha256:6895744dd558485f63c5c3c5f0b6df5fd4b376f364f0bd7feee554d889c121b4

Observation 70139650-2f18-4ab0-b2de-5dfa75baba63 · outbound

This paper cites arXiv preprint arXiv:2603.00945 , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions arXiv preprint arXiv:2603.00945 , year =

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-15T14:39:17.144034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.238646Z digest=sha256:f2775f38613bf29a15591acca84ddb88995515c12b5c2bd8b0d53de1f697e114

Observation 9f37f0ef-f166-4472-8cb7-778a0fd55dfe · outbound

This paper cites Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning , year =

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.241863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.241863Z digest=sha256:571ec0c4c9539327cbd962741ae30e6915e6b3379172eb7f47c27e009a9abd14

Observation 3e7d40ca-995e-44ca-9dc9-ed77e017b67b · outbound

This paper cites Efficiently Solving.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Efficiently Solving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.244721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.244721Z digest=sha256:32b02c49d886db66014a370b3cedf9ea22701ba44a8bcb10c1cfc7451b326e3a

Observation ae41af26-e4af-445c-aae9-695e441f1fc5 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Twelfth International Conference on Learning Representations , year =

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.247625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.247625Z digest=sha256:02a74748af1cab5ea666bb106fc8463f6e33dd0df66395663a828c58d2ce7582

Observation 76f441f9-afbb-4805-87b0-bc8a917a1e15 · outbound

This paper cites The Plugin Approach for Average-Reward and Discounted.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Plugin Approach for Average-Reward and Discounted

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.250801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.250801Z digest=sha256:db14d56658e480f8820b8d2311240355ba9afd9ba8d3cf83cf5cf3f90a95d852

Observation 88f0423b-98b0-4821-8345-22477419d669 · outbound

This paper cites Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.253403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.253403Z digest=sha256:dca22f8d69250ee09e92e62ef18d3fd07621f7bdd8cb60af0602ce340124448b

Observation effe7f4f-ba8d-490d-9f9b-38e263cabdf7 · outbound

This paper cites Sharper Model-Free Reinforcement Learning for Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sharper Model-Free Reinforcement Learning for Average-Reward

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.256314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.256314Z digest=sha256:d5c8ecfb7893df73bd3bd6a7c68d0ba031609a89be737dc3bc2cfd1ebf85ad6e

Observation bc7cecdf-4059-422f-8f9e-50a6354b90d3 · outbound

This paper cites Journal of Machine Learning Research , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Journal of Machine Learning Research , volume =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.259013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.259013Z digest=sha256:491311ddcecc0cb984b97ac67cbe4b58c7dab5558127f5cbce1298d0aa2ab7d9

Observation 5866a8ca-c052-40bf-bc7f-c07cdd22a80d · outbound

This paper cites Model-free Reinforcement Learning in Infinite-horizon Average-reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Model-free Reinforcement Learning in Infinite-horizon Average-reward

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.261581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.261581Z digest=sha256:09caa0a9be0de0b76c5b099847c9d025ce5aec1225aba4e43db5fb29c8974ff8

Observation bda1fb27-dbc1-44e4-8275-9035f35c413e · outbound

This paper cites Learning Infinite-horizon Average-reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Learning Infinite-horizon Average-reward

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.264641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.264641Z digest=sha256:a045b0b4c8fd675e8cf0e6c499ab4981d7cb75dab6bf0c79fd7a680e882caf0b

Observation 52a1e67a-ab1f-4cdf-8bfd-44fdaeb2fcea · outbound

This paper cites Efficient.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Efficient

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.267560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.267560Z digest=sha256:1d962b267362886fa4f03b69a18295d6329c7a984ad4a1ac8884f89f6d4fe06e

Observation 060f7f15-ac12-4f1e-a64a-c3b93494a9ab · outbound

This paper cites Distributionally Robust.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally Robust

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.270443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.270443Z digest=sha256:542482cf9a064e78577e81f4bb1abd68645c05361b48874c6e70bd254f6f4a50

Observation bb522753-d61b-4193-9151-66965ad1ffe4 · outbound

This paper cites The International Journal of Robotics Research , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The International Journal of Robotics Research , volume =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.273226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.273226Z digest=sha256:23834db3631174476e910bb49a49b5fee9c300d335a829108a6c8c5c660635f4

Observation df27e9a6-4869-4efa-af6f-0ab6da554678 · outbound

This paper cites Nature , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Nature , volume =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.276155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.276155Z digest=sha256:c72358d44108f585c676b513c753253f000282dce7f98936197b6955293de327

Observation 7f30be99-0561-4d7a-b0ef-3ab61d57ac2c · outbound

This paper cites Mastering the Game of.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mastering the Game of

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.279523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.279523Z digest=sha256:d32aa68c3d6b7b888ecec74d273940fc6f7801fb5dfbf691c6ba328242d04181

Observation cdd71e41-3f14-437e-a229-b55c90a318e1 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Artificial Intelligence and Statistics , pages =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.282444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.282444Z digest=sha256:8cbced37a6e7160397c039cd4aa648dc29d1643469cc8ec96ff79e23bd5b77b9

Observation 2b7b49e0-2026-41da-9bd1-bff29f14ac5c · outbound

This paper cites 2020 , organization =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions 2020 , organization =

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.285478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.285478Z digest=sha256:6470f6aff1f762ee5c0f0f1bcaaca1ab7b8d58141b746ebcd4c825e9c4837394

Observation d48defea-6d63-4673-a321-6bd36d16a06d · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.288546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.288546Z digest=sha256:7afe9c2e05a179f5b91fdd7b32948d11b1c42b3f759fb4cce3515995274c2c2e

Observation f24d8ef9-b189-4daa-b706-59e90a279f67 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Artificial Intelligence and Statistics , pages =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.291628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.291628Z digest=sha256:d9671fd41cbe033d13e21703e63b1a01b14f07e632a2df9e27f28ac5c3fd4089

Observation ab448d6c-4b6e-4cca-84fc-118f4a32ca85 · outbound

This paper cites COLT 2009 - The 22nd Conference on Learning Theory , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions COLT 2009 - The 22nd Conference on Learning Theory , year =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.294253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.294253Z digest=sha256:0e154c364886501267e6ad62b30a8b1bae9182bb6de7e421d057a6ef5d5699ed

Observation caf66f2d-f67c-4a94-88f2-c20bd87974ba · outbound

This paper cites Data-Driven Distributionally Robust Optimization Using the.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Data-Driven Distributionally Robust Optimization Using the

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.297291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.297291Z digest=sha256:4cc8a234d78c24debd92248f657aff551c0086023bfd57ffe05ac5997d2e3183

Observation dadd3c77-37ac-4010-a482-bbc94ff2f375 · outbound

This paper cites Distributionally robust convex optimization , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally robust convex optimization , year =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.300073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.300073Z digest=sha256:7d9a0d96df80d12656e22c9e56d52817d02f2f5fd948dac37da780f8c80c483c

Observation 8d85ffe9-de8a-4cd5-99c2-9f2c346e662b · outbound

This paper cites Distributionally robust optimization and its tractable approximations , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally robust optimization and its tractable approximations , year =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.303083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.303083Z digest=sha256:5b44a08bd2e77926e5afea1388d6d066747eaf4f261d072bc752d5c0c1986f66

Observation 5066e41a-5384-44dc-aa61-aa46226213b7 · outbound

This paper cites Learning models with uniform performance via distributionally robust optimization , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Learning models with uniform performance via distributionally robust optimization , year =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.306435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.306435Z digest=sha256:b3bbbe97f72e3d349bc49924f12453468405a36ac8007de11c89fdbd9a1a5f99

Observation ef0a72be-43a9-46dd-b59e-e6da2f75aab8 · outbound

This paper cites Robust Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Robust Average-Reward

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.309479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.309479Z digest=sha256:8127a0aac5db11f94e3e96943a7eb4810f8f7df1b0b01f10351b2107c9277cee

Observation 13d7ae5b-197e-4f50-99e9-e2f79fbf928d · outbound

This paper cites and Prater-Bennette, Ashley and Zou, Shaofeng , booktitle =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions and Prater-Bennette, Ashley and Zou, Shaofeng , booktitle =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.312817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.312817Z digest=sha256:ccfdbeb5eb63c250082ec6b9f3ab83f428bc74762eb97886bec92b98dd7f81b7

Observation c85eb2ac-fcb5-4b2a-969e-a4e561107ce5 · outbound

This paper cites Toward Theoretical Understandings of Robust.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Toward Theoretical Understandings of Robust

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.315761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.315761Z digest=sha256:1fc547adb4318e566408ae66a64c4ff06314559ddb44bff15567cf0581cd480b

Observation c9a23339-85dc-413f-8167-7dc904fef751 · outbound

This paper cites Sample Complexity of Variance-Reduced Distributionally Robust.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sample Complexity of Variance-Reduced Distributionally Robust

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.318707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.318707Z digest=sha256:56b9aacd604d08151bc55cddcde8600415601544ff4dc7745ee6355e752e1c76

Observation 5b71829d-561e-49c5-9191-50ddc86959cf · outbound

This paper cites Near-Optimal Distributionally Robust Reinforcement Learning with General.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near-Optimal Distributionally Robust Reinforcement Learning with General

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.322481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.322481Z digest=sha256:beb601ddca14284a9f4c4ea3fe706958cdf57c0590860b4ee071b361c9de7a7c

Observation 51e9ba48-0ed9-4ecc-8bc8-9d3e96119b5e · outbound

This paper cites Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity , year =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.325515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.325515Z digest=sha256:50f0bf269985016efea1db4ecbab381f729469db9b167c0eff97d5fa4d928f9d

Observation 82caf014-8a07-4f00-a1b6-b4f5deb0e0e0 · outbound

This paper cites Sample Complexity of Offline Distributionally Robust Linear.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sample Complexity of Offline Distributionally Robust Linear

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.328595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.328595Z digest=sha256:ca4bb42a279e314745f6535a52e77c11e81b5bed54e8e084313a67c4d62b506b

Observation 0277484c-a2c1-4750-8320-96faf0343c2f · outbound

This paper cites 1994 , publisher =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions 1994 , publisher =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.331697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.331697Z digest=sha256:0ad27c02d445f361e06de910249c79554599b365d24621e19df51814ae336d2c

Observation 0acdd3e8-5b79-46af-849f-79700514230d · outbound

This paper cites Tsybakov , publisher =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Tsybakov , publisher =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.334697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.334697Z digest=sha256:105dfc8efb0413cc37e4b8a82c7b986615e906b13ef4b2e13a61b21c9cd85956

Observation f9b615a3-36c1-47b2-a5fc-34953045dd77 · outbound

This paper cites Machine learning , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Machine learning , volume=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.338841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.338841Z digest=sha256:844afc551453e791ab80b8c859f7d09dcee0f43db3d65150029a90b5730f3d75

Observation d42131d3-33cc-4460-a376-4f7a909dfe1e · outbound

This paper cites International Conference on Machine Learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Machine Learning , pages=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.341917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.341917Z digest=sha256:925ba678994e374241d2bbc1b0e6ac5b19aca40e0147d223f3fd725ba47217fa

Observation 219b76fe-3d4c-467a-afbe-9f27746d6e13 · outbound

This paper cites Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.344469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.344469Z digest=sha256:5ee00b3adf1fa5de2f9773e63a563f8d9c5c12164c843160b05cbeda1493fc60

Observation d613f426-39b7-478e-b447-630fadb22f2e · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Forty-second International Conference on Machine Learning , year=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.347703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.347703Z digest=sha256:61bae166c54b960e905034d86be3f2d4224dd58e5b40acf78032f1da01686b4f

Observation 9babee25-36ec-45a5-8cd0-f2e18149d86e · outbound

This paper cites The blessing of heterogeneity in federated.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The blessing of heterogeneity in federated

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.350738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.350738Z digest=sha256:77c524e008c8f3783bc9229d2786ce95a754f9c5a92aac64e85153481462a980

Observation 3d5003de-2530-4f7f-a5e2-59e8f1b38eed · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.353842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.353842Z digest=sha256:2784f45575d94b81ece0123b71360fde0c148cace64ba1db045c075150b5f3ef

Observation 3b7dc32c-427a-4975-b155-89302f03eebd · outbound

This paper cites Thirty-seventh Conference on Neural Information Processing Systems , year=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Thirty-seventh Conference on Neural Information Processing Systems , year=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.356422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.356422Z digest=sha256:10d28bb907e3fae08ca5d1d60757e6294c83dc734c7e8fc0b5c430012398c61c

Observation d03d0526-5a80-4b81-b762-847536a57436 · outbound

This paper cites Operations research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Operations research , volume=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.359852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.359852Z digest=sha256:878394e71d043e473ba4d96d2a757fee12a19d71571064adaaef45fb62c590b2

Observation a3b7a8dd-08e4-44e2-9442-3a14b31ff866 · outbound

This paper cites Robust control of.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Robust control of

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.486678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.486678Z digest=sha256:f3f0536f8da6644661b429cc9d56ac021725b4565e0189f23c135f2a5ec606a4

Observation cdfc98d3-e4d7-4f31-bdef-0eeadf2f775e · outbound

This paper cites $Q$-learning with Logarithmic Regret.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions $Q$-learning with Logarithmic Regret

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:17.048249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.493614Z digest=sha256:d551a672d9f3fefb7f2b7d90d2ee11d8dd6f6028163fc7e69c0a9ab3d765124a

Observation 791b6d3f-07a0-4b23-b368-86e548e84c4f · outbound

This paper cites Near-Optimal Provable Uniform Convergence in Offline Policy Evaluation for Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near-Optimal Provable Uniform Convergence in Offline Policy Evaluation for Reinforcement Learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.496949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.496949Z digest=sha256:f28d58510601c1c013c9d27477f6c27f726a12c20e5b283fcc4b114734a36432

Observation 78457264-6cba-4039-9fd9-2f848b71ca06 · outbound

This paper cites Advances in neural information processing systems , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , pages=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.500102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.500102Z digest=sha256:fb919494c1a2353835442206c465715754316e04026bbc84e07cbcd55f94e9f3

Observation 4d4d9a1c-2280-4df3-8905-23cfc31a8ac3 · outbound

This paper cites Complete Dictionary Learning via $\ell_p$-norm Maximization.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Complete Dictionary Learning via $\ell_p$-norm Maximization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.503037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.503037Z digest=sha256:844ee3bc8b7f608cb58336d6e155003a7792e316ed50c578c3df4c8807ce22cc

Observation 4c2ba3de-423a-48a3-993f-8e7b0fc6da5f · outbound

This paper cites Journal of Applied Probability , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Journal of Applied Probability , volume=

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.506776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.506776Z digest=sha256:8c6bda08ed2c3077686a7006ba8c6fda82bd39ca6145e0edc0d9619d393b1373

Observation d88536b2-f16f-4da6-955f-bbe6e9d0198e · outbound

This paper cites Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.509987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.509987Z digest=sha256:0d6264bcb42ee91a905d7ec6f49939e101bfd4ed13237b456338eb203a09d0db

Observation b285e844-8231-4ac4-b382-a11c23509a4c · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.513098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.513098Z digest=sha256:edf3a9f01fae614e8c16fa59db9f3bc55298353e5e66bb2100c5ae4fdd6f931a

Observation a2746f34-f148-4e9d-b074-5d3f586fae8e · outbound

This paper cites On the Global Convergence Rates of Softmax Policy Gradient Methods.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions On the Global Convergence Rates of Softmax Policy Gradient Methods

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.516180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.516180Z digest=sha256:6563df888af64258287be146ea065ae92c04c8d8eaeec63deb8d211e16881b2e

Observation 9bfa37da-e391-4468-aad1-830e70ba60a6 · outbound

This paper cites ICML , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions ICML , volume=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.519221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.519221Z digest=sha256:259dea8a3405795a9a1ff744ca2781bcfae9692261deaabcf5c172aed3cabaab

Observation 24b203a2-f37d-4dad-9454-be3634036b70 · outbound

This paper cites Optimality and approximation with policy gradient methods in.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Optimality and approximation with policy gradient methods in

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.521951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.521951Z digest=sha256:619737f8f6d333d8d8b2e9777889cee06f04d72b7a0782d86a3999f119ba8e96

Observation 4619700e-2526-4d70-b16f-2937434744ee · outbound

This paper cites Advances in neural information processing systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , volume=

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.524694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.524694Z digest=sha256:26341324baa161790798bcc96cc2e900d33cc88f295495edd0d43b11b44740ee

Observation 6acaff08-1837-4d96-8ec8-579d4069fbd3 · outbound

This paper cites Advances in neural information processing systems , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , pages=

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.527449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.527449Z digest=sha256:864869114af57653837b210a381f52ba079d2e0a4d14cacffa03b02b4c99ba21

Observation bf8226aa-05b0-4c66-bf10-6907e51e48cc · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.530633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.530633Z digest=sha256:5ec3275088d9f0c97cbc7d62f704ceec3ae6a80d75c55fa5d6729b3ce6383cbb

Observation bb02d565-1b38-4e88-ba39-9677c0ba4def · outbound

This paper cites Nearly Minimax Optimal Regret for Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Nearly Minimax Optimal Regret for Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:16.986137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.533537Z digest=sha256:28e8c66e45155591176132e2be08ccfaca02cc31544408bbf9be6143eaf3b8b0

Observation 41ab3e61-0120-460f-adce-2841f65c14c5 · outbound

This paper cites Probability Theory and Related Fields , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Probability Theory and Related Fields , volume=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.536769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.536769Z digest=sha256:28b534fb2642f47579440a1c7fce7b156d8738404c42091670c09ce885b865a5

Observation ad91d121-3c2d-4722-8a4f-b3625fa945fd · outbound

This paper cites Proceedings of the 27th international conference on international conference on machine learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Proceedings of the 27th international conference on international conference on machine learning , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.540035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.540035Z digest=sha256:3f2194d463b609b6f35fa686b36f427d364e9f578d6931494d937c76018c88dc

Observation 61a978c1-dc40-42d6-8a93-e2f6ea1fb7fa · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Learning Representations (ICLR) , year=

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.543051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.543051Z digest=sha256:cfd487a9f18bf3d605e9f78526c6e2fca081f0653eb8ddf605245d4d492915d2

Observation 5a63e9ef-e709-4519-a2f2-b938816c1d4e · outbound

This paper cites Theoretical Linear Convergence of Unfolded ISTA and its Practical Weights and Thresholds.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Theoretical Linear Convergence of Unfolded ISTA and its Practical Weights and Thresholds

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.545887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.545887Z digest=sha256:b40d7439b3467d9422370880d5f3672959e3894fd161fbd77c578297b32ca69c

Observation 835631c6-af5a-4d43-8270-834dab3ff1ba · outbound

This paper cites Ada-LISTA: Learned Solvers Adaptive to Varying Models.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Ada-LISTA: Learned Solvers Adaptive to Varying Models

Reference 71

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:16.960039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.548922Z digest=sha256:242636b481ba0c274c11c22f6073d9bce6911b9bb1f08482488cf5fc5e23f009

Observation caac32fe-da59-4a36-a272-c0efd64b4a3c · outbound

This paper cites Understanding Trainable Sparse Coding via Matrix Factorization.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Understanding Trainable Sparse Coding via Matrix Factorization

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.552169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.552169Z digest=sha256:189274ba6ae1ba950e8152e113236ace645e82d40e2683bab59d51ada37c7498

Observation 14528f54-1f2e-4b30-ab9b-7c9c74523b67 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Artificial Intelligence and Statistics , pages=

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.555278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.555278Z digest=sha256:be57650679ca3e325d8ee5c5d44cf0d48b3a4222900190ef829b11d62e38ffd1

Observation bc2f6af1-29a9-45fd-b665-3ca9bf0dc9fe · outbound

This paper cites Mathematics of Operations Research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mathematics of Operations Research , volume=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.558053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.558053Z digest=sha256:e8209be92455034b7053ba82c81f283115e673d18b6dd41a7e57dbf8c16a361b

Observation 3dde4795-b613-4951-b572-7d0e891208ea · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume=

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.560606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.560606Z digest=sha256:fd205aeabc49d58e0696debadc5b98aaaa48cd43515c9dd35cfefc3c530db31a

Observation 0e6eac17-e3d8-47c5-96cb-ecc2c259c858 · outbound

This paper cites Twice regularized.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Twice regularized

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.563461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.563461Z digest=sha256:2deac52111c2af208a771264ec0604305f15aca89f0c4056def72c7d0a753ac8

Observation 5e06f2a4-1d0e-41aa-b21b-c3ef8b779bcb · outbound

This paper cites International Conference on Machine Learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Machine Learning , pages=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.566048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.566048Z digest=sha256:5c75936f46c7d02553cc6bc9f7270975a2c627e9fc56023a22c4a68f46c7d987

Observation 1ce8b195-a019-46e1-a093-05fdf0f79678 · outbound

This paper cites A Review of Off-Policy Evaluation in Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions A Review of Off-Policy Evaluation in Reinforcement Learning

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.568605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.568605Z digest=sha256:5b86c09c61671d1a6dba1d50556d23a54c6c2c79c2bb364ecd96e5105a8859a1

Observation 73c7a7ad-9d56-440f-a888-20da053818af · outbound

This paper cites International Conference on Machine Learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Machine Learning , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.571404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.571404Z digest=sha256:51b1f1f0154bd79796962d61144cebac9fd1ac18167550d82c51f42721de571e

Observation 1fd2fac0-ba16-429b-aad4-5d1ca99d20bc · outbound

This paper cites Distributionally Robust Optimization: A Review.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally Robust Optimization: A Review

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.574358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.574358Z digest=sha256:7199ffc516a019804b79563fb6b1ad31683459f283a4d6fd20544bc7958bff67

Observation 95b26c2f-0f01-4095-a4a2-0ece99f3e556 · outbound

This paper cites Finite-sample guarantees for.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Finite-sample guarantees for

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.578729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.578729Z digest=sha256:0a7d13afda83ab91a8747ccf0a7fca0a5e988e2ddb4bacf142330b451aed60a2

Observation 11359dc3-2b4b-499a-9887-c87535093a98 · outbound

This paper cites Certifying Model Accuracy under Distribution Shifts.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Certifying Model Accuracy under Distribution Shifts

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.581826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.581826Z digest=sha256:d120c31b4c13ee93212b1ec23a2e5d35c79db1674818be2cdeee1f357aeea107

Observation a76b1b01-4c57-423d-a0bf-bc2a6f1facfb · outbound

This paper cites Advances in neural information processing systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , volume=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.585147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.585147Z digest=sha256:09c0258f7d0c269cfe38057cccf73c97e1c8488322cb2faa9d013ab4d2e7e9ed

Observation 3fd88426-9fe4-4008-b4d0-9a108e0bdd80 · outbound

This paper cites Settling the Sample Complexity of Model-Based Offline Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Settling the Sample Complexity of Model-Based Offline Reinforcement Learning

Reference 84

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:16.893053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.588609Z digest=sha256:7c23bb3ac2056cfaa303067c7ef64823afe890d74dbf39b6db18bb6627d9e4fb

Observation d72da13f-f3f1-4232-974a-19cff24db1e6 · outbound

This paper cites Available at Optimization Online , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Available at Optimization Online , pages=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.591949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.591949Z digest=sha256:59962878c99d32f02fa452d6befdf8ff7dea43e45af62379fd44e432dd44f6f9

Observation 81a5dc8b-44de-43f4-97e9-b1e3e9a9078a · outbound

This paper cites Pessimistic.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Pessimistic

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.594947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.594947Z digest=sha256:4868f25649c1c3bdc153d6236af3dd51f1b6a18aee5cc92b6385a1f7edcd8142

Observation 1ca62b82-0e54-4ad5-94c9-e6599cd5f221 · outbound

This paper cites The Bell system technical journal , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Bell system technical journal , volume=

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.597974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.597974Z digest=sha256:d7082fd45e1855155d041a28e7514b17799d19904cf2e82d4d85bb5a5ff10a9b

Observation 3bf77a69-8377-415e-9085-5610c6dd79f5 · outbound

This paper cites The mathematics of data , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The mathematics of data , volume=

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.600564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.600564Z digest=sha256:41d653d9599f082977a68a24bf726030c31a2b060ab888c8f796f5526729fe0b

Observation 0e2948c7-0976-4bd5-a85a-bce669fd91ee · outbound

This paper cites The Journal of finance , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Journal of finance , volume=

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.603762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.603762Z digest=sha256:c4d6021a50cbf3fb5e1157dbc7cb0c8d7ed4de73938764a207031e078cd87c99

Observation a06183ee-4d5e-455b-a185-f794e2047b6b · outbound

This paper cites , author=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions , author=

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.606758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.606758Z digest=sha256:4ee5bee9eb3155450ca88125bfa52ab5faae4d24dc10f63d85c4075d3fd1d2ac

Observation 47828712-199e-4990-a716-666c81d0da09 · outbound

This paper cites Mathematical Programming , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mathematical Programming , volume=

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.609707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.609707Z digest=sha256:ef06ebefd9af95c341a9121ce327e496cebddc715b2afdc1d54a4ff4d8cebab0

Observation df1e378e-fe72-4bd8-b904-22dd29d668a4 · outbound

This paper cites Mathematics of Operations Research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mathematics of Operations Research , volume=

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.613255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.613255Z digest=sha256:d0a54841dd541ea46becf0f84f2da0199e23ef224099f37d1c520dd61b44d29a

Observation b86932a2-17df-4343-9aa1-13f973b37687 · outbound

This paper cites Robust control of uncertain.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Robust control of uncertain

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.616496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.616496Z digest=sha256:071d8303e3c7baa9ee2066dd0ee0beefbbfa34430b6a04c16bd92c4250f5673d

Observation a521443e-c233-48e4-876b-650f335feb2a · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.619844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.619844Z digest=sha256:d0b3eae73b2987e7cc17c31918cad9768569db14a49ddef50c2e8a6eac431088

Observation d86bc39b-ddcb-44e3-8a07-023b3e422c4d · outbound

This paper cites INFORMS Journal on Computing , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions INFORMS Journal on Computing , volume=

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.622784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.622784Z digest=sha256:8bd7a6da3c7386f27d8d7a025ebbf72628670edf7ec22611164cf5bfbec02b73

Observation c290e883-4b9c-42ec-b04f-c5ed1283589c · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.625757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.625757Z digest=sha256:68f31a77e4dc5dd6aba8115b6b947d4149f3504618b8219291735d684330f67b

Observation 29f6aba9-0e16-4f56-af81-020f27f8653c · outbound

This paper cites Distributionally Robust Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally Robust Reinforcement Learning

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.628743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.628743Z digest=sha256:f50bc486756c8c739a839fb1f6596cb82f502145862ac1379e01de7691f50ef6

Observation e8b0c55a-7f0a-445a-a4ca-a48965dd18c3 · outbound

This paper cites Journal of Machine Learning Research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Journal of Machine Learning Research , volume=

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.631794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.631794Z digest=sha256:179a68076f184a6d541e4049a447005b303a6c73996825fa962db3eed8ee08fa

Observation a6bed617-be62-448b-b66d-d30b9ee729a8 · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.634562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.634562Z digest=sha256:64028615072f7fd8b7fa22aba80025663829e91a7a554f14b87495fef25f9803

Observation 88f1c433-0416-48a4-bdf1-f5f8f2a8cef7 · outbound

This paper cites Distributional Robustness and Regularization in Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributional Robustness and Regularization in Reinforcement Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.637632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.637632Z digest=sha256:ab4eca2310a9ec6e06f782db6c671e6502d739e01d7c9982de90a30ae3417a7a

Pith citing papers

No inbound Pith citation observations are available.