首页 /研究 /Importance sampling for online planning under uncertainty
OTHER

Importance sampling for online planning under uncertainty

Yuanfu Luo, Haoyu Bai, David Hsu, Wee Sun Lee

发表年份
2018
引用次数
49

摘要

The partially observable Markov decision process (POMDP) provides a principled general framework for robot planning under uncertainty. Leveraging the idea of Monte Carlo sampling, recent POMDP planning algorithms have scaled up to various challenging robotic tasks, including, real-time online planning for autonomous vehicles. To further improve online planning performance, this paper presents IS-DESPOT, which introduces importance sampling to DESPOT, a state-of-the-art sampling-based POMDP algorithm for planning under uncertainty. Importance sampling improves DESPOT’s performance when there are critical, but rare events, which are difficult to sample. We prove that IS-DESPOT retains the theoretical guarantee of DESPOT. We demonstrate empirically that importance sampling significantly improves the performance of online POMDP planning for suitable tasks. We also present a general method for learning the importance sampling distribution.

关键词

Partially observable Markov decision processSampling (signal processing)Computer scienceArtificial intelligenceMachine learningMarkov chain Monte CarloSample (material)Markov chainBayesian probabilityComputer vision

相关论文

查看 OTHER 分类全部论文