首页 /研究 /Model-based policy search for automatic tuning of multivariate PID controllers
LEARNING

Model-based policy search for automatic tuning of multivariate PID controllers

Andreas Doerr, Duy Nguyen-Tuong, Alonso Marco, Stefan Schaal, Sebastian Trimpe

发表年份
2017
引用次数
3

摘要

PID control architectures are widely used in industrial applications. Despite their low number of open parameters, tuning multiple, coupled PID controllers can become tedious in practice. In this paper, we extend PILCO, a model-based policy search framework, to automatically tune multivariate PID controllers purely based on data observed on an otherwise unknown system. The system's state is extended appropriately to frame the PID policy as a static state feedback policy. This renders PID tuning possible as the solution of a finite horizon optimal control problem without further a priori knowledge. The framework is applied to the task of balancing an inverted pendulum on a seven degree-of-freedom robotic arm, thereby demonstrating its capabilities of fast and data-efficient policy learning, even on complex real world problems.

关键词

PID controllerInverted pendulumControl theory (sociology)Computer scienceA priori and a posterioriControl engineeringControl (management)State (computer science)Frame (networking)Multivariable calculus

相关论文

查看 LEARNING 分类全部论文