Convergent Reinforcement Learning with Value Function Interpolation

Abstract

We consider the convergence of a class of reinforcement learning algorithms combined with value function interpolation methods using the methods developed in (Littman and Szepesvari, 1996). As a special case of the obtained general results, for the first time, we prove the (almost sure) convergence of Q-learning when combined with value function interpolation in uncountable spaces.

BibTeX key: szepesvari2000
entry type: techreport
address: Budapest 1121, Konkoly Th. M. u. 29-33, HUNGARY
year: 2000
institution: Mindmaker Ltd.
number: TR-2001-02
pdf: papers/rlfapp.pdf
date-modified: 2010-09-04 14:48:33 -0600

Users

Comments and Reviewsshow / hide

Please log in to take part in the discussion (add own reviews or comments).

BibSonomy

Convergent Reinforcement Learning with Value Function Interpolation

Abstract

Tags

Users

Comments and Reviewsshow / hide

Cite this publication

More citation styles

search on