Metric learning for reinforcement learning agents
Matthew E. Taylor, Brian Kulis, Fei Sha
- 发表年份
- 2011
- 引用次数
- 19
摘要
A key component of any reinforcement learning algorithm is the underlying representation used by the agent. While reinforcement learning (RL) agents have typically relied on hand-coded state rep-resentations, there has been a growing interest in learning this rep-resentation. While inputs to an agent are typically fixed (i.e., state variables represent sensors on a robot), it is desirable to automati-cally determine the optimal relative scaling of such inputs, as well as to diminish the impact of irrelevant features. This work intro-duces HOLLER, a novel distance metric learning algorithm, and combines it with an existing instance-based RL algorithm to achieve precisely these goals. The algorithms ’ success is highlighted via empirical measurements on a set of six tasks within the mountain car domain.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002