Published June 13, 2020 | Version Accepted Version
Discussion Paper Open

The Power of Predictions in Online Control

Abstract

We study the impact of predictions in online Linear Quadratic Regulator control with both stochastic and adversarial disturbances in the dynamics. In both settings, we characterize the optimal policy and derive tight bounds on the minimum cost and dynamic regret. Perhaps surprisingly, our analysis shows that the conventional greedy MPC approach is a near-optimal policy in both stochastic and adversarial settings. Specifically, for length-T problems, MPC requires only O(logT) predictions to reach O(1) dynamic regret, which matches (up to lower-order terms) our lower bound on the required prediction horizon for constant regret.

Additional Information

This project was supported in part by funding from Raytheon, DARPA PAI, AitF-1637598 and CNS-1518941, with additional support for Guanya Shi provided by the Simoudis Discovery Prize. We see no ethical concerns related to the results in this paper.

Attached Files

Accepted Version - 2006.07569.pdf

Files

2006.07569.pdf

Files (586.6 kB)

Name Size
md5:63cd5abdeb76f88300ab4a207b216fc1
586.6 kB Preview Download

Additional details

Identifiers

Eprint ID
104236
Resolver ID
CaltechAUTHORS:20200707-094715120

Funding

Raytheon Company
Defense Advanced Research Projects Agency (DARPA)
NSF
CCF-1637598
NSF
CNS-1518941
Simoudis Discovery Prize

Dates

Created
2020-07-07
Created from EPrint's datestamp field
Updated
2023-06-02
Created from EPrint's last_modified field

Caltech Custom Metadata

Caltech groups
GALCIT