Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:35:55.558707Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 2 inbound Pith citation observations for arXiv:2505.09518.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:35:55.558707Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:35:55.428858Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-15T21:35:55.594608Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 73c46f7d-ab06-4136-8156-c654c4db1fec · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Scalable internal-state policy-gradient methods for POMDPs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b309826d-20dd-4f0c-b720-e3802eb345ac · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 54f8d2e9-253e-49ba-b2c1-f1bd9ebd7915 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Pessimistic Iterative Planning with RNNs for Robust POMDPs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4f1e4b0-197e-4d8f-805c-35af3cd8efd0 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fdffc7df-441a-4006-b79e-4f86c5a63121 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Multi-cost bounded tradeoff analysis in MDP
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation acbd0d91-c95b-4aa8-bb34-e110b6fa32dd · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Gradient- descent for randomized controllers under partial observ- ability
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8e228d68-c9b8-4377-96f1-716ed4cfa995 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Littman, and Anthony R
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f16f666b-263a-463e-b027-d6be7e8b4c54 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs PRISM 4.0: Verification of proba- bilistic real-time systems
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c0719c59-751f-4806-b6ac-2f4ef02f74c8 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Littman, Anthony R
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2afd1a78-2b6c-4e13-9282-d82d62f7b4dd · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Policy iteration for bounded-parameter POMDPs
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b3373df1-1fb0-4750-b6a8-0f2b1ef8532c · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Multiple-environment Markov decision processes
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6bb30a1b-84c7-4788-9bbe-0e9955df26e5 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Ambiguous partially observable Markov decision processes: Structural results and applications
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ee28828a-7701-4466-b203-0e3a24e381d9 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Domain randomization for transferring deep neural net- works from simulation to the real world
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 635c7391-c1c8-47fb-97fb-8f765d5f07a6 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Policy gradient in robust MDPs with global conver- gence guarantee
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2199623b-4390-408d-ae62-fb72a05987cb · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Robust Markov decision processes.Math
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a77906f6-49cb-40d2-ae34-6185b5d699e1 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Cost-bounded active clas- sification using partially observable markov decision pro- cesses
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0aa496db-49a8-4ee1-8c36-e8751e2ef1a4 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs We use a fixed step size, i.e.,αk =α> 0, that we set toα = 0.1
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f7e40471-6a8c-4bf8-90de-f6ebc439a72e · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Stochastic modeling of a power-managed system: con- struction and optimization
Reference 1994
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 29433022-aeac-45ba-9bd3-e40182b9ce10 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs On the undecidability of probabilistic planning and infinite-horizon partially observable Markov decision problems
Reference 1997
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a94b119d-b5fa-446d-a28c-bfc4a981b2e3 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Levy, and Shie Mannor
Reference 1998
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f9929093-2342-43da-b161-ec1c7fb46a85 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Distributionally robust partially observable Markov decision process with moment-based ambiguity
Reference 1999
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0870e8a6-0801-4f53-b300-40646940fd5b · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Cassandra
Reference 2000
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3aa614ca-a362-4a61-ba66-925446c2c672 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Policy-gradient algo- rithms for partially observable Markov decision processes
Reference 2002
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a45b2a0f-65c1-4320-a2d4-7234c880afd7 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Finite-state controllers based on Mealy machines for centralized and decentralized POMDPs
Reference 2003
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a6a331c3-c2f6-4d4b-9c77-01f151b980ee · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Finite-state controllers of POMDPs using parameter synthesis
Reference 2005
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0e78958a-a330-4e4c-af78-e92818e33fc4 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs PAYNT: a tool for inductive synthe- sis of probabilistic programs
Reference 2010
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c41401a3-11c9-462c-9254-1d4bdec9368f · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs A single-loop robust policy gradient method for robust Markov decision processes
Reference 2011
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cef39d61-a77c-4c28-924c-2993cffce751 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Optimal cost almost-sure reachability in POMDPs
Reference 2012
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dfcdf329-076c-4bc3-9d46-51ac04aebe00 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Verification and control of partially observable probabilistic systems
Reference 2013
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f720be2c-4868-4610-86c7-d0f1c043f6e1 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Learning robust policies for uncertain parametric Markov decision processes
Reference 2014
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f0f3057b-cc8b-4aba-b49e-99520b72bf46 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Puterman
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d968f439-f515-4844-952a-5f6b7c84f780 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Multiple-environment Markov decision processes: Efficient analysis and applications
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 472ca84d-a8b2-4d1a-88df-bedc2ed73440 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Robust partially ob- servable Markov decision process
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 986a5416-adb1-40ee-8150-10be8c43f907 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Robust finite-state controllers for uncertain POMDPs
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 51da2508-8529-4c60-8ea3-b85a9fbb8237 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Real-time scheduling over Markovian chan- nels: When partial observability meets hard deadlines
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8eaeb912-d071-4c11-a312-8a36324d9805 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs ProFeat: feature- oriented engineering for family-based probabilistic model checking
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 71e68da2-1b65-4016-bc78-8ba157052437 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Search and explore: Symbiotic policy synthesis in POMDPs
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 64a519a1-5790-430f-93bc-09bd0a4879bb · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3acbbc78-ee9e-4968-ae5f-047ddc7a2a42 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Policies grow on trees: Model checking families of MDPs
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bf7825fb-6186-4d00-a306-f6bc8fcbfa61 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Martin, Samuel Nicol, R´egis Sabbadin, and Olivier Buffet
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 90f01092-de56-427e-9cea-1bca99a6d7c8 · outbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Scalable first-order methods for robust MDPs
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a4f1e4b0-197e-4d8f-805c-35af3cd8efd0 · inbound
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 79d01f92-8e09-4a1d-a044-a73562782aeb · inbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.