首页 /研究 /Supervised ActorCritic Reinforcement Learning
LEARNING

Supervised ActorCritic Reinforcement Learning

发表年份
2009
引用次数
50

摘要

Chapter 7 introduced policy gradients as a way to improve on stochastic search of the policy space when learning. This chapter presents supervised actor-critic reinforcement learning as another method for improving the effectiveness of learning. With this approach, a supervisor adds structure to a learning problem and supervised learning makes that structure part of an actor-critic framework for reinforcement learning. Theoretical background and a detailed algorithm description are provided, along with several examples that contain enough detail to make them easy to understand and possible to duplicate. These examples also illustrate the use of two kinds of supervisors: a feedback controller that is easily designed yet suboptimal, and a human operator providing intermittent control of a simulated robotic arm.

关键词

Reinforcement learningComputer scienceArtificial intelligenceSupervisorMachine learningSupervised learningError-driven learningLearning classifier systemActive learning (machine learning)Semi-supervised learning

相关论文

查看 LEARNING 分类全部论文