pith. machine review for the scientific record. sign in

arxiv: 1207.4166 · v1 · submitted 2012-07-11 · 💻 cs.AI

Recognition: unknown

Heuristic Search Value Iteration for POMDPs

Authors on Pith no claims yet
classification 💻 cs.AI
keywords hsvivalueiterationpomdpsearchalgorithmheuristicliterature
0
0 comments X
read the original abstract

We present a novel POMDP planning algorithm called heuristic search value iteration (HSVI).HSVI is an anytime algorithm that returns a policy and a provable bound on its regret with respect to the optimal policy. HSVI gets its power by combining two well-known techniques: attention-focusing search heuristics and piecewise linear convex representations of the value function. HSVI's soundness and convergence have been proven. On some benchmark problems from the literature, HSVI displays speedups of greater than 100 with respect to other state-of-the-art POMDP value iteration algorithms. We also apply HSVI to a new rover exploration problem 10 times larger than most POMDP problems in the literature.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Synthesizing POMDP Policies: Sampling Meets Model-checking via Learning

    cs.AI 2026-05 conditional novelty 6.0

    A new synthesis framework for POMDPs learns finite-state controllers via sampling and model-checking oracles, achieving relative completeness when the policy is regular and solving threshold-safety problems beyond exi...