English

Meta-Reinforcement Learning for the Tuning of PI Controllers: An Offline Approach

Systems and Control 2022-09-20 v2 Machine Learning Systems and Control

Abstract

Meta-learning is a branch of machine learning which trains neural network models to synthesize a wide variety of data in order to rapidly solve new problems. In process control, many systems have similar and well-understood dynamics, which suggests it is feasible to create a generalizable controller through meta-learning. In this work, we formulate a meta reinforcement learning (meta-RL) control strategy that can be used to tune proportional--integral controllers. Our meta-RL agent has a recurrent structure that accumulates "context" to learn a system's dynamics through a hidden state variable in closed-loop. This architecture enables the agent to automatically adapt to changes in the process dynamics. In tests reported here, the meta-RL agent was trained entirely offline on first order plus time delay systems, and produced excellent results on novel systems drawn from the same distribution of process dynamics used for training. A key design element is the ability to leverage model-based information offline during training in simulated environments while maintaining a model-free policy structure for interacting with novel processes where there is uncertainty regarding the true process dynamics. Meta-learning is a promising approach for constructing sample-efficient intelligent controllers.

Keywords

Cite

@article{arxiv.2203.09661,
  title  = {Meta-Reinforcement Learning for the Tuning of PI Controllers: An Offline Approach},
  author = {Daniel G. McClement and Nathan P. Lawrence and Johan U. Backstrom and Philip D. Loewen and Michael G. Forbes and R. Bhushan Gopaluni},
  journal= {arXiv preprint arXiv:2203.09661},
  year   = {2022}
}

Comments

23 pages; postprint