Pith. sign in

Paper Citation Record · LEDGER

Linear Mixture Distributionally Robust Markov Decision Processes

As of 17 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2505.18044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18044 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:16.530086Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy43
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dc186288-6cb8-4ced-95f5-554b82e2a0c9 · outbound

This paper cites Improved algorithms for linear stochastic bandits.Advances in Neural Information Processing Systems, 24, 2011.

Linear Mixture Distributionally Robust Markov Decision Processes Improved algorithms for linear stochastic bandits.Advances in Neural Information Processing Systems, 24, 2011

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:27.385792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.019563Z digest=sha256:225bd02462e2983f405b6a0ae08dcc817a6388143f6a3352c484e65c9bb75ea2

Observation bb835027-f4f9-4f05-93ca-afff8eca1798 · outbound

This paper cites Model-based rein- forcement learning with value-targeted regression.

Linear Mixture Distributionally Robust Markov Decision Processes Model-based rein- forcement learning with value-targeted regression

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:27.195035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.070077Z digest=sha256:5b0bf08a48574433a1585ff641936213de8a6a59d5b78458624f94c2ff108505

Observation 38a13ecd-1e1d-4364-8365-9ae7148b5933 · outbound

This paper cites Robust reinforcement learning using least squares policy iteration with provable performance guarantees.

Linear Mixture Distributionally Robust Markov Decision Processes Robust reinforcement learning using least squares policy iteration with provable performance guarantees

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:26.981351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.135324Z digest=sha256:5d4ab5461351fd9489550ec3463fecf68a9c136f80ef46179fe6c3587878aa0a

Observation 00956383-db3a-4030-bb26-9b74abc7e6f1 · outbound

This paper cites an unresolved cited work.

Linear Mixture Distributionally Robust Markov Decision Processes Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:26.789328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.254486Z digest=sha256:e0d68dc372e28f462c312bfbf7934b09ed07a3ac838d4a6c1071e918294cc7a2

Observation 1dd62c58-047c-4ffc-8b5b-baf919b0cf50 · outbound

This paper cites Provably efficient exploration in policy optimization.

Linear Mixture Distributionally Robust Markov Decision Processes Provably efficient exploration in policy optimization

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:26.541681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.322725Z digest=sha256:5370f4b12d62c45769db360ed50274f5baa6a00e754efe713fab61fbfc866f1c

Observation 3f20851a-ba71-43a7-af1d-095e670e23a8 · outbound

This paper cites A Survey of Sim-to-Real Methods in RL: Progress, Prospects and Challenges with Foundation Models.

Linear Mixture Distributionally Robust Markov Decision Processes A Survey of Sim-to-Real Methods in RL: Progress, Prospects and Challenges with Foundation Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:10.442723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:10.442723Z digest=sha256:54db067626a60f4819bace83e300e09a2ae99ccfb73ae7dfcc7360387002ad58

Observation 19a16ae0-0857-41d3-ae1b-e2da5c47ff67 · outbound

This paper cites Off-dynamics reinforcement learning: Training for transfer with domain classifiers.

Linear Mixture Distributionally Robust Markov Decision Processes Off-dynamics reinforcement learning: Training for transfer with domain classifiers

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:26.293774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.544737Z digest=sha256:41d4819176e959ec8abd74a9054202ba90889949389652ea85f32e987075017c

Observation 54623bc2-edaa-4cb8-993d-1db9562a20cc · outbound

This paper cites Birkhauser Boston Inc., 1989.

Linear Mixture Distributionally Robust Markov Decision Processes Birkhauser Boston Inc., 1989

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.988409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.646720Z digest=sha256:575f7e4b341e801d95e050903867d2f1bbe10ffc0af529f8ec49787fb8c3e685

Observation 7ab5c91c-f782-4b02-9065-0bc0a743a499 · outbound

This paper cites Robust markov decision processes: Beyond rectangu- larity.Mathematics of Operations Research, 48(1):203–226, 2023.

Linear Mixture Distributionally Robust Markov Decision Processes Robust markov decision processes: Beyond rectangu- larity.Mathematics of Operations Research, 48(1):203–226, 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.708718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.741252Z digest=sha256:ef1bfa82f76a38d1ee8d90e400155389d75a95e1382182201ca0ec01c8ef1516

Observation 3fa3af11-b595-4805-931b-4d939d270587 · outbound

This paper cites Off-dynamics reinforcement learning via domain adaptation and reward augmented imitation.

Linear Mixture Distributionally Robust Markov Decision Processes Off-dynamics reinforcement learning via domain adaptation and reward augmented imitation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.473463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.836792Z digest=sha256:01fbbcb5190cb42704c20e717df94c05d66152f8d4ed93c517d44514d2be4f5c

Observation 8cf0cc27-d3e2-42cc-94fc-00e9988764f3 · outbound

This paper cites Kullback-leibler divergence constrained distributionally robust optimization.Available at Optimization Online, 1(2):9, 2013.

Linear Mixture Distributionally Robust Markov Decision Processes Kullback-leibler divergence constrained distributionally robust optimization.Available at Optimization Online, 1(2):9, 2013

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:25.114752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:10.941541Z digest=sha256:fddd7832a05ca6cf3a6da130213d4d24808822bfd4ddc42d6f3b22259934849b

Observation 26f5d310-fae8-4dec-9df9-918a18143f24 · outbound

This paper cites Robust dynamic programming.Mathematics of Operations Research, 30(2): 257–280, 2005.

Linear Mixture Distributionally Robust Markov Decision Processes Robust dynamic programming.Mathematics of Operations Research, 30(2): 257–280, 2005

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.825016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.093164Z digest=sha256:d581f1e263d24b7e3cb94bda8b22d23826b50f1b309e372f41c5cb7539ee9ab0

Observation 1133a7aa-aac7-4269-9920-365395c561b4 · outbound

This paper cites Model-based reinforcement learning with value-targeted regression.

Linear Mixture Distributionally Robust Markov Decision Processes Model-based reinforcement learning with value-targeted regression

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.634036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.207448Z digest=sha256:00ca658d7d8c5b6104bd096d604fef4a9a0b3c72dc1457b268ac6edb35933d47

Observation 9681b5ac-3886-4ad9-a7e8-705f2973edb4 · outbound

This paper cites Reinforcement learning in robotics: A survey.

Linear Mixture Distributionally Robust Markov Decision Processes Reinforcement learning in robotics: A survey

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.415638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.315068Z digest=sha256:3bfdb8310e9d1b95f917d83504c61ad233ded6cf60fc7efd174c51f361b3bb49

Observation a8f9b543-73ed-47b2-be81-feb4db80c716 · outbound

This paper cites The transferability approach: Crossing the reality gap in evolutionary robotics.IEEE Transactions on Evolutionary Computa- tion, 17(1):122–145, 2012.

Linear Mixture Distributionally Robust Markov Decision Processes The transferability approach: Crossing the reality gap in evolutionary robotics.IEEE Transactions on Evolutionary Computa- tion, 17(1):122–145, 2012

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:24.211608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.487182Z digest=sha256:17dcc4ea99b459a007a9d5183424372a8d09bed1c96fe4c656db1cf436d641ca

Observation 537e3031-7b53-4ac4-84f6-ac234e66cc16 · outbound

This paper cites Improved algorithm for adversarial linear mixture mdps with bandit feedback and unknown transition.

Linear Mixture Distributionally Robust Markov Decision Processes Improved algorithm for adversarial linear mixture mdps with bandit feedback and unknown transition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.992506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.642883Z digest=sha256:0371476ff511395fada022eb013857a5ae1f4618b12e08f2c77db180bfeff275

Observation 9e4b0ec0-dad0-4e2b-8215-312c05ec1e7f · outbound

This paper cites Policy gradient algorithms for robust mdps with non-rectangular uncertainty sets.arXiv preprint arXiv:2305.19004, 2023.

Linear Mixture Distributionally Robust Markov Decision Processes Policy gradient algorithms for robust mdps with non-rectangular uncertainty sets.arXiv preprint arXiv:2305.19004, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:11.789167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:11.789167Z digest=sha256:cff68332b4e3d82dac211085630942a488b9c8fe6a9799857e1b657a588aaa50

Observation c0a9d7d5-2e5b-490a-9ce1-c4a8d716093d · outbound

This paper cites Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.793901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.853527Z digest=sha256:b49b26a625636401f0ac0dab95acc320abb4b70b18e6886f03ed99d07c95c085

Observation 0dc062d7-014f-4736-813b-841708978fd2 · outbound

This paper cites Minimax optimal and computationally efficient algorithms for distributionally robust offline reinforcement learning.

Linear Mixture Distributionally Robust Markov Decision Processes Minimax optimal and computationally efficient algorithms for distributionally robust offline reinforcement learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.553624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:11.967284Z digest=sha256:3abb6aa2afa010b847d7fa0589096b9b5ba6ed8da2dd6662415d21b7acfe5347

Observation 63daded1-ac5d-49ab-99aa-5b63af275ba0 · outbound

This paper cites Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning.

Linear Mixture Distributionally Robust Markov Decision Processes Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.083195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:12.083195Z digest=sha256:864a4c8d67659956eaac0d3825aa928f94ff7992ff6b6848e2f86f70bf9b541c

Observation 95e6eb6c-acaa-4371-9d67-69ea426d175c · outbound

This paper cites Distributionally robust reinforcement learning with interactive data collection: Fundamental hardness and near-optimal algorithms.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally robust reinforcement learning with interactive data collection: Fundamental hardness and near-optimal algorithms

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.385864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.199629Z digest=sha256:02422e076a1d1297285282c9be3b565d34c08dd398d6b3cea097826f55c411a1

Observation b85144d0-96dd-4589-b3c1-7a46c22f8c9f · outbound

This paper cites Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.293596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:12.293596Z digest=sha256:136a3938e65f12fbe423ae52366237502bf3c76f40d172a0c5f55c86224f9757

Observation 0c3ffa72-4d05-4171-ad13-ea9788ec4a94 · outbound

This paper cites Finite mixture models.Annual review of statistics and its application, 6(1):355–378, 2019.

Linear Mixture Distributionally Robust Markov Decision Processes Finite mixture models.Annual review of statistics and its application, 6(1):355–378, 2019

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.288836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.397984Z digest=sha256:f353378a00ee9e28285fafbdec47e0feb4faa1b8addc8bf6dab142f22a0ab908

Observation 44bb74d5-08e3-42dd-ae8a-af04cf923cc2 · outbound

This paper cites A simplex method for function minimization.The computer journal, 7(4):308–313, 1965.

Linear Mixture Distributionally Robust Markov Decision Processes A simplex method for function minimization.The computer journal, 7(4):308–313, 1965

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.223487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.495205Z digest=sha256:a0093b6fa58f50b7b9e8cdb396677563a75a2add06f3bb996f7e7f26ec3590d3

Observation 67835c7f-4f71-4a3e-97e5-8b9fc42752ed · outbound

This paper cites Robust control of markov decision processes with uncertain transition matrices.Operations Research, 53(5):780–798, 2005.

Linear Mixture Distributionally Robust Markov Decision Processes Robust control of markov decision processes with uncertain transition matrices.Operations Research, 53(5):780–798, 2005

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.169848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.579883Z digest=sha256:0e33c64a459fe549d341c1b2fc9a7c0a4aa3bd5b5331b8958af6f560edc984ec

Observation 839e7291-ab3a-497e-8e24-f86eb3b3790a · outbound

This paper cites Robustness in markov decision problems with uncertain transition matrices.Advances in Neural Information Processing Systems, 16, 2003.

Linear Mixture Distributionally Robust Markov Decision Processes Robustness in markov decision problems with uncertain transition matrices.Advances in Neural Information Processing Systems, 16, 2003

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:23.014790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.671049Z digest=sha256:1cd5da34584d0c0615651adad25ef9370e477f38257a592b888f174ce15ad12c

Observation 18d66843-81be-44cc-aa87-48dd55c772f0 · outbound

This paper cites Assessing Generalization in Deep Reinforcement Learning.

Linear Mixture Distributionally Robust Markov Decision Processes Assessing Generalization in Deep Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.736776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:12.736776Z digest=sha256:9bce7526ea1c88131b5d49b2d34f9da6b3e7d0e00aecc9ac0eb7745c54c75934

Observation 4142db4e-26e7-4179-b4b9-ce97031e44e5 · outbound

This paper cites Bridging distributionally robust learning and offline rl: An approach to mitigate distribution shift and partial data coverage.

Linear Mixture Distributionally Robust Markov Decision Processes Bridging distributionally robust learning and offline rl: An approach to mitigate distribution shift and partial data coverage

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.701706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.810675Z digest=sha256:97ccc0698306d3b81fdb78afa9261a76572085a666811add3fd534b296c9658f

Observation e8fd1073-752e-49a0-92fe-9ccba380a9ab · outbound

This paper cites The infinite gaussian mixture model.Advances in Neural Information Processing Systems, 12, 1999.

Linear Mixture Distributionally Robust Markov Decision Processes The infinite gaussian mixture model.Advances in Neural Information Processing Systems, 12, 1999

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.533967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.922041Z digest=sha256:f9ca42beae7b7cf28bc4f7f6060df736904489169e82051bfc8e58f9f17b74ef

Observation 45337095-b368-4d61-abe8-bec4486f2fed · outbound

This paper cites Gaussian mixture models.Encyclopedia of biometrics, 741(659-663): 3, 2009.

Linear Mixture Distributionally Robust Markov Decision Processes Gaussian mixture models.Encyclopedia of biometrics, 741(659-663): 3, 2009

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.342941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:12.982231Z digest=sha256:3c2e7c712775a4d10cdbaa7dde17aeb3bf1d0e1bbf69b8ae070bc054a6245c27

Observation e64b59e7-8913-4b22-a6da-19a224eb9698 · outbound

This paper cites Markovian decision processes with uncertain transition probabilities.Operations Research, 21(3):728–740, 1973.

Linear Mixture Distributionally Robust Markov Decision Processes Markovian decision processes with uncertain transition probabilities.Operations Research, 21(3):728–740, 1973

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.165116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:13.089462Z digest=sha256:d785ba19927fee2d03eee080121eb15bbe4ec5a781b1a87684ffbeb33f94b729

Observation 942aebe0-2956-4962-aff9-008cd03e2f88 · outbound

This paper cites Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity.Journal of Machine Learning Research, 25(200):1–91,.

Linear Mixture Distributionally Robust Markov Decision Processes Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity.Journal of Machine Learning Research, 25(200):1–91,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:22.019192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:13.224367Z digest=sha256:0f028fc714816190981beca2fa506008e3d9abcabb2184056cb6d7c2bd8eb39c

Observation ac69338f-c324-45ef-b6b7-b78ecf2e27b3 · outbound

This paper cites The curious price of distributional robustness in reinforcement learning with a generative model.Advances in Neural Information Processing Systems, 36, 2024.

Linear Mixture Distributionally Robust Markov Decision Processes The curious price of distributional robustness in reinforcement learning with a generative model.Advances in Neural Information Processing Systems, 36, 2024

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.820136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:13.360333Z digest=sha256:43fd47e1b75fc144eb6b615ac1112fb36ca8e69c94b2c4af672b42381aefcdb5

Observation 1bc9241f-01ad-4e82-9451-126cbac39427 · outbound

This paper cites Robust offline reinforcement learning with linearly structuredf-divergence regularization.arXiv preprint arXiv:2411.18612, 2024.

Linear Mixture Distributionally Robust Markov Decision Processes Robust offline reinforcement learning with linearly structuredf-divergence regularization.arXiv preprint arXiv:2411.18612, 2024

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.540609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:13.540609Z digest=sha256:386901c2b4884e0fcbdd0bdfa83e9f1d46cfe4959ed88809277a4e904fef365e

Observation b1cf2ff8-fe45-4251-b08c-791f50ab0925 · outbound

This paper cites Pessimistic model-based offline reinforcement learning under partial coverage.

Linear Mixture Distributionally Robust Markov Decision Processes Pessimistic model-based offline reinforcement learning under partial coverage

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.615475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:13.759936Z digest=sha256:a29377870ceb9f9e9fb52cf77d494c1d3a91a6ad5483bd927e7e0b8ad5a2b481

Observation 3e8d53b8-5bfb-4d31-816b-e06f9bef2976 · outbound

This paper cites Sample complexity of offline distributionally robust linear markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes Sample complexity of offline distributionally robust linear markov decision processes

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.435252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:13.955582Z digest=sha256:03da4a2afaf7ed519bd1d8f060dff5fa524a5c45d4475553dd42aad7728b411f

Observation 48e2cbd0-62f7-4b33-aa0d-c0db5afc2550 · outbound

This paper cites Return augmented decision transformer for off-dynamics reinforcement learning.arXiv preprint arXiv:2410.23450, 2024.

Linear Mixture Distributionally Robust Markov Decision Processes Return augmented decision transformer for off-dynamics reinforcement learning.arXiv preprint arXiv:2410.23450, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:14.180938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:44:14.180938Z digest=sha256:80d98e17576aaa19f759efd85f714059ba985d7b1f0b408496d16d0d8fdaec36

Observation e85e83e1-f41d-410e-990b-32bc20e4b662 · outbound

This paper cites Robust markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes Robust markov decision processes

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.282437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:14.366644Z digest=sha256:15c33591cba1949a5c7bbb25eb9863538c1611b12f0d282ad436eef933e007e5

Observation 5cd48174-e78c-4762-9c2e-2fc32074bb3d · outbound

This paper cites Mutual alignment transfer learning.

Linear Mixture Distributionally Robust Markov Decision Processes Mutual alignment transfer learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:21.148819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:14.512261Z digest=sha256:bc6b941bb933a99d43e96531fc8952060b6579554f9373d60ed4f10cc74d4622

Observation f019cdc2-eb5f-4a4f-b270-8c78b598182f · outbound

This paper cites The robustness-performance tradeoff in markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes The robustness-performance tradeoff in markov decision processes

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.851485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:14.695106Z digest=sha256:f0f30d87e531a4869b36fe304455acba0d0ae5a7ba94ac37a8602980b3a059fd

Observation 4c789318-dc3a-46c2-942a-78328245af70 · outbound

This paper cites Improved sample complexity bounds for distributionally robust reinforcement learning.

Linear Mixture Distributionally Robust Markov Decision Processes Improved sample complexity bounds for distributionally robust reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.560793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:14.854406Z digest=sha256:347a117c41400a704487585d1f1c57b5fe2a31e2970663bf0d9b216a17bdd24a

Observation 48fe2c9b-761c-4683-95c9-cbe1ca36377a · outbound

This paper cites Toward theoretical understandings of robust markov decision processes: Sample complexity and asymptotics.The Annals of Statistics, 50 (6):3223–3248, 2022.

Linear Mixture Distributionally Robust Markov Decision Processes Toward theoretical understandings of robust markov decision processes: Sample complexity and asymptotics.The Annals of Statistics, 50 (6):3223–3248, 2022

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.325566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:15.061034Z digest=sha256:6dc230c8f07d717acf2accdf1055dabcbaa6a7dfa7577b72a64ad25adf7f774e

Observation 9ea4557c-3d4f-4c06-8821-0998a738ab90 · outbound

This paper cites Reward-free model-based reinforcement learning with linear function approximation.Advances in Neural Information Processing Systems, 34:1582–1593, 2021.

Linear Mixture Distributionally Robust Markov Decision Processes Reward-free model-based reinforcement learning with linear function approximation.Advances in Neural Information Processing Systems, 34:1582–1593, 2021

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:20.071385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:15.218794Z digest=sha256:377d832c3cca6593ba5faee5e2a4838936eb93a6dedaaebb1d144da7927560b9

Observation c9bd51f5-0c82-4c63-b83e-3b55b28179df · outbound

This paper cites Learning adversarial linear mixture markov decision processes with bandit feedback and unknown transition.

Linear Mixture Distributionally Robust Markov Decision Processes Learning adversarial linear mixture markov decision processes with bandit feedback and unknown transition

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:19.715374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:15.452113Z digest=sha256:7f841801df526a61ce15bb70255587ae7e1de9165ff5bcf43f95e6dd0c93db58

Observation 64543293-787a-4e7d-9421-ba5098cc83b2 · outbound

This paper cites Sim-to-real transfer in deep reinforcement learning for robotics: a survey.

Linear Mixture Distributionally Robust Markov Decision Processes Sim-to-real transfer in deep reinforcement learning for robotics: a survey

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:19.344128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:15.598520Z digest=sha256:ff5442d77a0c65cc6afeb63511acb34ea04284d3e8596d5224de4455f8903efd

Observation c31db8ea-24ae-41f8-9c98-2f9ac1ce89a9 · outbound

This paper cites Nearly minimax optimal reinforcement learning for linear mixture markov decision processes.

Linear Mixture Distributionally Robust Markov Decision Processes Nearly minimax optimal reinforcement learning for linear mixture markov decision processes

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:18.984048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:15.708272Z digest=sha256:d806e1607337b8723873932f93cfeb7ce623a56df30e6a26c33a7e0111ca1aaa

Observation e07a3f5e-17ce-4ad3-bd72-8388f5f6787a · outbound

This paper cites Provably efficient reinforcement learning for discounted mdps with feature mapping.

Linear Mixture Distributionally Robust Markov Decision Processes Provably efficient reinforcement learning for discounted mdps with feature mapping

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:18.616950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:15.877799Z digest=sha256:99612bf329a78badfb4a780d7e7b5db385f8f1ea1c00e4833fb78af95739f274

Observation 749cdfc5-2079-4a84-9f5c-0edfc4defa36 · outbound

This paper cites Natural actor-critic for robust reinforcement learning with function approximation.

Linear Mixture Distributionally Robust Markov Decision Processes Natural actor-critic for robust reinforcement learning with function approximation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:18.244415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:16.093379Z digest=sha256:3ed3918b948324a0061be2f18111bf791334ccfe8f63aebc72fc8960717f4c90

Observation e0055a81-ecff-4ada-b7a5-12dacbbec727 · outbound

This paper cites Finite-sample regret bound for distributionally robust offline tabular reinforcement learning.

Linear Mixture Distributionally Robust Markov Decision Processes Finite-sample regret bound for distributionally robust offline tabular reinforcement learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:17.858529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:16.285333Z digest=sha256:87d33f56ce99a1772de5bfc907da77b4f3937fda28bdb555e562fc33ee5c1f61

Observation 832614ab-83b0-4b22-b97b-5e9906b73a12 · outbound

This paper cites Time- constrained robust mdps.Advances in Neural Information Processing Systems, 37:35574–35611,.

Linear Mixture Distributionally Robust Markov Decision Processes Time- constrained robust mdps.Advances in Neural Information Processing Systems, 37:35574–35611,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:17.527348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:16.401146Z digest=sha256:6793ed899254becf9a58606a411c78751e63ab98c660840f44d38188098c5f58

Observation 5a065462-4302-4047-8c1a-f3386aceef79 · outbound

This paper cites A.1 Proof of Theorem 3.4 Proof.

Linear Mixture Distributionally Robust Markov Decision Processes A.1 Proof of Theorem 3.4 Proof

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:17.145142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:44:16.530086Z digest=sha256:703fe912f2486ad209060afe69e36ac7c4e497bb2e1a60ddee867f67153735c8

Pith citing papers

No inbound Pith citation observations are available.