首页 /研究 /Optimal Control of Markov Decision Processes With Linear Temporal Logic Constraints
OTHER

Optimal Control of Markov Decision Processes With Linear Temporal Logic Constraints

Xuchu Ding, Stephen L. Smith, Călin Belta, Daniela Rus

发表年份
2014
引用次数
179

摘要

In this paper, we develop a method to automatically generate a control policy for a dynamical system modeled as a Markov Decision Process (MDP). The control specification is given as a Linear Temporal Logic (LTL) formula over a set of propositions defined on the states of the MDP. Motivated by robotic applications requiring persistent tasks, such as environmental monitoring and data gathering, we synthesize a control policy that minimizes the expected cost between satisfying instances of a particular proposition over all policies that maximize the probability of satisfying the given LTL specification. Our approach is based on the definition of a novel optimization problem that extends the existing average cost per stage problem. We propose a sufficient condition for a policy to be optimal, and develop a dynamic programming algorithm that synthesizes a policy that is optimal for a set of LTL specifications.

关键词

Markov decision processLinear temporal logicTemporal logicComputer scienceMathematical optimizationPartially observable Markov decision processDynamic programmingMarkov processOptimal controlSet (abstract data type)

相关论文

查看 OTHER 分类全部论文