Model-Based Policy Search for Automatic Tuning of Multivariate PID
  Controllers

Doerr, Andreas; Marco, Alonso; Nguyen-Tuong, Duy; Schaal, Stefan; Trimpe, Sebastian

research

Model-Based Policy Search for Automatic Tuning of Multivariate PID Controllers

Authors: Andreas Doerr
Alonso Marco
Duy Nguyen-Tuong
Stefan Schaal
Sebastian Trimpe
Publication date: 8 March 2017
Publisher
Doi

Abstract

PID control architectures are widely used in industrial applications. Despite their low number of open parameters, tuning multiple, coupled PID controllers can become tedious in practice. In this paper, we extend PILCO, a model-based policy search framework, to automatically tune multivariate PID controllers purely based on data observed on an otherwise unknown system. The system's state is extended appropriately to frame the PID policy as a static state feedback policy. This renders PID tuning possible as the solution of a finite horizon optimal control problem without further a priori knowledge. The framework is applied to the task of balancing an inverted pendulum on a seven degree-of-freedom robotic arm, thereby demonstrating its capabilities of fast and data-efficient policy learning, even on complex real world problems.Comment: Accepted final version to appear in 2017 IEEE International Conference on Robotics and Automation (ICRA

Similar works

Full text

Available Versions

Crossref

info:doi/10.1109%2Ficra.2017.7...

Last time updated on 06/08/2021