首页 /研究 /Min-Max Entropy Inverse RL of Multiple Tasks
OTHER

Min-Max Entropy Inverse RL of Multiple Tasks

Saurabh Arora, Prashant Doshi, Bikramjit Banerjee

发表年份
2021
引用次数
2

摘要

Multi-task IRL recognizes that expert(s) could be switching between multiple ways of solving the same problem, or interleaving demonstrations of multiple tasks. The learner aims to learn the reward functions that individually guide these distinct ways. We present a new method for multi-task IRL that generalizes the well-known maximum entropy approach by combining it with a Dirichlet process based minimum entropy clustering of the observed data. This yields a single nonlinear optimization problem, called MinMaxEnt Multi-task IRL (MME-MTIRL), which can be solved using the Lagrangian relaxation and gradient descent methods. We evaluate MME-MTIRL on the robotic task of sorting onions on a processing line where the expert utilizes multiple ways of detecting and removing blemished onions. The method is able to learn the underlying reward functions to a high level of accuracy and it improves on the previous approaches.

关键词

Computer scienceInterleavingCluster analysisArtificial intelligencePrinciple of maximum entropyEntropy (arrow of time)

相关论文

查看 OTHER 分类全部论文