Conferences >2008 IEEE International Confe...

Bayesian reinforcement learning in continuous POMDPs with application to robot navigation

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

We consider the problem of optimal control in continuous and partially observable environments when the parameters of the model are not known exactly. Partially observabl...Show More

Metadata

Abstract:

We consider the problem of optimal control in continuous and partially observable environments when the parameters of the model are not known exactly. Partially observable Markov decision processes (POMDPs) provide a rich mathematical model to handle such environments but require a known model to be solved by most approaches. This is a limitation in practice as the exact model parameters are often difficult to specify exactly. We adopt a Bayesian approach where a posterior distribution over the model parameters is maintained and updated through experience with the environment. We propose a particle filter algorithm to maintain the posterior distribution and an online planning algorithm, based on trajectory sampling, to plan the best action to perform under the current posterior. The resulting approach selects control actions which optimally trade-off between 1) exploring the environment to learn the model, 2) identifying the system's state, and 3) exploiting its knowledge in order to maximize long-term rewards. Our preliminary results on a simulated robot navigation problem show that our approach is able to learn good models of the sensors and actuators, and performs as well as if it had the true model.

Published in: 2008 IEEE International Conference on Robotics and Automation

Date of Conference: 19-23 May 2008

Date Added to IEEE Xplore: 13 June 2008

ISBN Information:

Print ISSN: 1050-4729

DOI: 10.1109/ROBOT.2008.4543641

Conference Location: Pasadena, CA, USA

Contents

References is not available for this document.

Bayesian reinforcement learning in continuous POMDPs with application to robot navigation

Abstract:

Metadata

Abstract:

References

IEEE Account

Purchase Details

Profile Information

Need Help?

Bayesian reinforcement learning in continuous POMDPs with application to robot navigation

Alerts

Abstract:

Metadata

Abstract:

References

IEEE Account

Purchase Details

Profile Information

Need Help?