Minimax Q-learning Control for Linear Systems Using the Wasserstein Metric

Feiran Zhao,Keyou You
DOI: https://doi.org/10.1016/j.automatica.2022.110850
IF: 6.4
2023-01-01
Automatica
Abstract:Stochastic optimal control usually requires an explicit dynamical model with probability distributions, which are difficult to obtain in practice. In this work, we consider the linear quadratic regulator (LQR) problem of unknown linear systems and adopt a Wasserstein penalty to address the distribution uncertainty of additive stochastic disturbances. By constructing an equivalent deterministic game of the penalized LQR problem, we propose a Q-learning method with convergence guarantees to learn an optimal minimax controller.
What problem does this paper attempt to address?