Monte-Carlo utility estimates for Bayesian reinforcement learning

Christos Dimitrakakis

doi:10.1109/CDC.2013.6761048

Monte-Carlo utility estimates for Bayesian reinforcement learning
Paper in proceeding, 2013

This paper introduces a set of algorithms for Monte-Carlo Bayesian reinforcement learning. Firstly, Monte-Carlo estimation of upper bounds on the Bayes-optimal value function is employed to construct an optimistic policy. Secondly, gradient-based algorithms for approximate upper and lower bounds are introduced. Finally, we introduce a new class of gradient algorithms for Bayesian Bellman error minimisation. We theoretically show that the gradient methods are sound. Experimentally, we demonstrate the superiority of the upper bound method in terms of reward obtained. However, we also show that the Bayesian Bellman error method is a close second, despite its significant computational simplicity.

Author

Christos Dimitrakakis

Chalmers, Computer Science and Engineering (Chalmers), Computing Science (Chalmers)

Other publications Research

Proceedings of the IEEE Conference on Decision and Control

07431546 (ISSN) 25762370 (eISSN)

7303-7308 6761048
978-1-4673-5717-3 (ISBN)

Areas of Advance

Information and Communication Technology

Subject Categories

Computational Mathematics

Probability Theory and Statistics

Control Engineering

Computer Science

DOI

10.1109/CDC.2013.6761048

Publication data connected to DOI

ISBN

978-1-4673-5717-3

More information

Latest update

1/3/2024 9

If you have questions, need help, find a bug or just want to give us feedback you may use this form, or contact us per e-mail research.lib@chalmers.se.

Message

Your email address

Research.chalmers.se contains research information from Chalmers University of Technology, Sweden. It includes information on projects, publications, research funders and collaborations.

More about coverage period and what is publicly available

Privacy and cookies

Accessibility

Citation Style Language
citeproc-js (Frank Bennett)

Chalmers Library

Chalmers Research

Chalmers Student Theses

SE-412 96 GOTHENBURG, SWEDEN
PHONE: +46 (0)31-772 10 00
WWW.CHALMERS.SE