Data-driven learning of feedback maps for explicit robust predictive control: an approximation theoretic view
Siddhartha Ganguly, Shubham Gupta, Debasish Chatterjee
TL;DR
The paper addresses robust MPC for uncertain linear systems by marrying data-driven learning with explicit policy synthesis. It solves the underlying minmax OCP exactly at grid points via a convex SIP solved with the MSA algorithm to generate state–action data, then learns explicit feedback maps with uniform error guarantees using QuIFS for low dimensions and NNFS for moderate dimensions. The authors prove that the learned policies preserve recursive feasibility and exhibit ISS-like stability, and they demonstrate two numerical examples showing improved region of attraction and fast online evaluation relative to traditional multiparametric methods. The approach offers a practical, provably reliable path to real-time robust control with data-driven explicit controllers, applicable to systems where online optimization is too costly.
Abstract
We establish an algorithm to learn feedback maps from data for a class of robust model predictive control (MPC) problems. The algorithm accounts for the approximation errors due to the learning directly at the synthesis stage, ensuring recursive feasibility by construction. The optimal control problem consists of a linear noisy dynamical system, a quadratic stage and quadratic terminal costs as the objective, and convex constraints on the state, control, and disturbance sequences; the control minimizes and the disturbance maximizes the objective. We proceed via two steps -- (a) Data generation: First, we reformulate the given minmax problem into a convex semi-infinite program and employ recently developed tools to solve it in an exact fashion on grid points of the state space to generate (state, action) data. (b) Learning approximate feedback maps: We employ a couple of approximation schemes that furnish tight approximations within preassigned uniform error bounds on the admissible state space to learn the unknown feedback policy. The stability of the closed-loop system under the approximate feedback policies is also guaranteed under a standard set of hypotheses. Two benchmark numerical examples are provided to illustrate the results.
