Pith. sign in

REVIEW 1 cited by

Policy iteration for perfect information stochastic mean payoff games with bounded first return times is strongly polynomial

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1310.4953 v1 pith:ODDNRQQF submitted 2013-10-18 math.OC

classification math.OC
keywords iterationpolicypolynomialstronglyboundeddiscountgamesrate
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Recent results of Ye and Hansen, Miltersen and Zwick show that policy iteration for one or two player (perfect information) zero-sum stochastic games, restricted to instances with a fixed discount rate, is strongly polynomial. We show that policy iteration for mean-payoff zero-sum stochastic games is also strongly polynomial when restricted to instances with bounded first mean return time to a given state. The proof is based on methods of nonlinear Perron-Frobenius theory, allowing us to reduce the mean-payoff problem to a discounted problem with state dependent discount rate. Our analysis also shows that policy iteration remains strongly polynomial for discounted problems in which the discount rate can be state dependent (and even negative) at certain states, provided that the spectral radii of the nonnegative matrices associated to all strategies are bounded from above by a fixed constant strictly less than 1.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Thresholds for sensitive optimality and Blackwell optimality in stochastic games

    cs.GT 2025-06 conditional novelty 7.0 of 10

    For perfect-information stochastic games, this paper derives the first explicit upper bounds on the d-sensitive discount threshold (for d≥0) and improved upper bounds on the Blackwell threshold.

Pith tools