SearcharxivSearch

arXiv subjects

Meriem Gharbi

Publications and source records attributed to Meriem Gharbi.

2 recordsLinked to original sources

Anytime Proximity Moving Horizon Estimation: Stability and Regret

In this paper, we address the efficient implementation of moving horizon state estimation of constrained discrete-time linear systems. We propose a novel iteration scheme which employs a proximity-based formulation of the underlying optimization algorithm and reduces computational effort by performing only a limited number of optimization iterations each time a new measurement is received. We outline conditions under which global exponential stability of the underlying estimation errors is ensured. Performance guarantees of the iteration scheme in terms of regret upper bounds are also established. A combined result shows that both exponential stability and a sublinear regret which can be rendered smaller by increasing the number of optimization iterations can be guaranteed. The stability and regret results of the proposed estimator are showcased through numerical simulations.

math.OC

Online learning with stability guarantees: A memory-based real-time model predictive controller

We propose and analyze a real-time model predictive control (MPC) scheme that utilizes stored data to improve its performance by learning the value function online with stability guarantees. For linear and nonlinear systems, a learning method is presented that makes use of basic analytic properties of the cost function and is proven to learn the MPC control law and the value function on the limit set of the closed-loop state trajectory. The main idea is to generate a smart warm start based on historical data that improves future data points and thus future warm starts. We show that these warm starts are asymptotically exact and converge to the solution of the MPC optimization problem. Thereby, the suboptimality of the applied control input resulting from the real-time requirements vanishes over time. Simulative examples show that existing real-time MPC schemes can be improved by storing data and the proposed learning scheme.

math.OC