Prioritization Methods for Accelerating MDP Solvers

David Wingate; Kevin D. Seppi

2005 JMLR JMLR 2005

Prioritization Methods for Accelerating MDP Solvers

Abstract

The performance of value and policy iteration can be dramatically improved by eliminating redundant or useless backups, and by backing up states in the right order. We study several methods designed to accelerate these iterative solvers, including prioritization, partitioning, and variable reordering. We generate a family of algorithms by combining several of the methods discussed, and present extensive empirical evidence demonstrating that performance can improve by several orders of magnitude for many problems, while preserving accuracy and convergence guarantees. [abs] [ pdf ][ bib ] © JMLR 2005. (edit, beta)

📈 Trend Setter — Value Iteration

🧭 Keyword Pioneer — value iteration

🐣 Hot Topic Early Bird — value iteration

🐝 Cross-Pollinator — Artificial Intelligence, Computer Vision, Deep Learning, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy

Authors

David Wingate , Kevin D. Seppi

Topics

Reinforcement Learning > Applications > Value Iteration

Keywords

value iteration policy iteration mdp solver state ordering

Download PDF

Related papers

Diffusion Kernels on Statistical Manifolds 2005

Learning with Decision Lists of Data-Dependent Features 2005

Multiclass Classification with Multi-Prototype Support Vector Machines 2005

Loopy Belief Propagation: Convergence and Effects of Message Errors 2005

Efficient Computation of Gapped Substring Kernels on Large Alphabets 2005