Representation Discovery for Kernel-Based Reinforcement Learning

dc.date.accessioned	2015-11-30T19:30:04Z
dc.date.accessioned	2018-11-26T22:27:30Z
dc.date.available	2015-11-30T19:30:04Z
dc.date.available	2018-11-26T22:27:30Z
dc.date.issued	2015-11-24
dc.identifier.uri	http://hdl.handle.net/1721.1/100053
dc.identifier.uri	http://repository.aust.edu.ng/xmlui/handle/1721.1/100053
dc.description.abstract	Recent years have seen increased interest in non-parametric reinforcement learning. There are now practical kernel-based algorithms for approximating value functions; however, kernel regression requires that the underlying function being approximated be smooth on its domain. Few problems of interest satisfy this requirement in their natural representation. In this paper we define Value-Consistent Pseudometric (VCPM), the distance function corresponding to a transformation of the domain into a space where the target function is maximally smooth and thus well-approximated by kernel regression. We then present DKBRL, an iterative batch RL algorithm interleaving steps of Kernel-Based Reinforcement Learning and distance metric adjustment. We evaluate its performance on Acrobot and PinBall, continuous-space reinforcement learning domains with discontinuous value functions.	en_US
dc.format.extent	16 p.	en_US
dc.rights	Creative Commons Attribution-ShareAlike 4.0 International
dc.rights.uri	http://creativecommons.org/licenses/by-sa/4.0/
dc.subject	Metric learning	en_US
dc.title	Representation Discovery for Kernel-Based Reinforcement Learning	en_US

Files in this item

Files	Size	Format	View
MIT-CSAIL-TR-2015-032.pdf	1.960Mb	application/pdf	View/Open

Except where otherwise noted, this item's license is described as Creative Commons Attribution-ShareAlike 4.0 International