Home /Research /Learning and planning with timing information in Markov decision processes
OTHER

Learning and planning with timing information in Markov decision processes

Pierre‐Luc Bacon, Borja Balle, Doina Precup

Year
2015
Citations
3

Abstract

We consider the problem of learning and planning in Markov decision processes with temporally extended actions represented in the options framework. We propose to use predictions about the duration of extended actions to represent the state and show that this leads to a compact predictive state representation model independent of the set of primitive actions. Then we develop a consistent and efficient spectral learning algorithm for such models. Using just the timing information to represent states allows for faster improvement in the planning performance. We illustrate our approach with experiments in both synthetic and robot navigation domains.

Keywords

Computer scienceMarkov decision processRepresentation (politics)Artificial intelligenceMachine learningMarkov chainMarkov processSet (abstract data type)State (computer science)Robot

Related papers

Browse all OTHER papers