首页 /研究 /Combination of learning from non-optimal demonstrations and feedbacks using inverse reinforcement learning and Bayesian policy improvement
LEARNING

Combination of learning from non-optimal demonstrations and feedbacks using inverse reinforcement learning and Bayesian policy improvement

Ali Ezzeddine, Nafee Mourad, Babak Nadjar Araabi, Majid Nili Ahmadabadi

发表年份
2018
引用次数
12

关键词

Computer scienceReinforcement learningTrainerTask (project management)Probabilistic logicArtificial intelligenceProcess (computing)Bayesian probabilityMachine learningHuman–computer interaction

相关论文

查看 LEARNING 分类全部论文