Current browse context:
stat.ML
Change to browse by:
References & Citations
Statistics > Machine Learning
Title: Optimal Reinforcement Learning for Gaussian Systems
(Submitted on 4 Jun 2011 (v1), last revised 14 Oct 2011 (this version, v3))
Abstract: The exploration-exploitation trade-off is among the central challenges of reinforcement learning. The optimal Bayesian solution is intractable in general. This paper studies to what extent analytic statements about optimal learning are possible if all beliefs are Gaussian processes. A first order approximation of learning of both loss and dynamics, for nonlinear, time-varying systems in continuous time and space, subject to a relatively weak restriction on the dynamics, is described by an infinite-dimensional partial differential equation. An approximate finite-dimensional projection gives an impression for how this result may be helpful.
Submission history
From: Philipp Hennig PhD [view email][v1] Sat, 4 Jun 2011 08:14:59 GMT (2456kb,AD)
[v2] Wed, 7 Sep 2011 16:11:15 GMT (37kb,D)
[v3] Fri, 14 Oct 2011 15:01:11 GMT (39kb,D)
Link back to: arXiv, form interface, contact.