Repository logo
Research Data
Publications
Projects
Persons
Organizations
English
Français
Log In(current)
  1. Home
  2. Publications
  3. Contribution à un congrès (conference paper)
  4. Monte-Carlo utility estimates for Bayesian reinforcement learning

Monte-Carlo utility estimates for Bayesian reinforcement learning

Author(s)
Dimitrakakis, Christos  
Chaire de science des données  
Date issued
2013
In
52nd IEEE Conference on Decision and Control
Subjects
Machine Learning (cs.LG) Machine Learning (stat.ML)
Abstract
This paper introduces a set of algorithms for Monte-Carlo Bayesian reinforcement learning. Firstly, Monte-Carlo estimation of upper bounds on the Bayes-optimal value function is employed to construct an optimistic policy. Secondly, gradient-based algorithms for approximate upper and lower bounds are introduced. Finally, we introduce a new class of gradient algorithms for Bayesian Bellman error minimisation. We theoretically show that the gradient methods are sound. Experimentally, we demonstrate the superiority of the upper bound method in terms of reward obtained. However, we also show that the Bayesian Bellman error method is a close second, despite its significant computational simplicity.
Publication type
conference paper
Identifiers
https://libra.unine.ch/handle/20.500.14713/21774
DOI
10.1109/CDC.2013.6761048
-
https://libra.unine.ch/handle/123456789/30979
File(s)
Loading...
Thumbnail Image
Download
Name

1303.2506.pdf

Type

Main Article

Size

131.22 KB

Format

Adobe PDF

Checksum

(MD5):414d6167751fdd838fb706b9d8119230

Université de Neuchâtel logo

Service information scientifique & bibliothèques

Rue Emile-Argand 11

2000 Neuchâtel

contact.libra@unine.ch

Service informatique et télématique

Rue Emile-Argand 11

Bâtiment B, rez-de-chaussée

Powered by DSpace-CRIS

v2.0.0

© 2025 Université de Neuchâtel

Portal overviewUser guideOpen Access strategyOpen Access directive Research at UniNE Open Access ORCIDWhat's new