Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

171

Voted

NIPS
2004

125views Information Technology» more NIPS 2004»

VDCBPI: an Approximate Scalable Algorithm for Large POMDPs

15 years 8 months ago

VDCBPI: an Approximate Scalable Algorithm for Large POMDPs

Download books.nips.cc

Existing algorithms for discrete partially observable Markov decision processes can at best solve problems of a few thousand states due to two important sources of intractability: the curse of dimensionality and the policy space complexity. This paper describes a new algorithm (VDCBPI) that mitigates both sources of intractability by combining the Value Directed Compression (VDC) technique [13] with Bounded Policy Iteration (BPI) [14]. The scalability of VDCBPI is demonstrated on synthetic network management problems with up to 33 million states.

Pascal Poupart, Craig Boutilier

Real-time Traffic

NIPS 2004 | NIPS 2007 | Observable Markov Decision | Policy Space Complexity | Value Directed Compression |

claim paper

Related Content

» Efficient Planning in Large POMDPs through Policy Graph Based Factorized Approximations

» Pointbased backup for decentralized POMDPs complexity and new algorithms

» Anytime PointBased Approximations for Large POMDPs

» POMDP Planning for Robust Robot Control

» Constraintbased dynamic programming for decentralized POMDPs with structured interactions

» Approximate Planning in POMDPs with MacroActions

» Theoretical Analysis of Heuristic Search Methods for Online POMDPs

» Exploiting domain knowledge in planning for uncertain robot systems modeled as POMDPs

» Pointbased value iteration An anytime algorithm for POMDPs

Post Info
More Details (n/a)

Added	31 Oct 2010
Updated	31 Oct 2010
Type	Conference
Year	2004
Where	NIPS
Authors	Pascal Poupart, Craig Boutilier

Comments (0)