CORTEXA
← Browse
arxiveess.SYmath.OC2026-07-01

A Data-Enabled Primal-Dual Approach for Policy Learning with SDP Formulations

Han Wang, Feiran Zhao, Florian Dorfler

This paper develops a data-enabled primal-dual framework for learning optimal control policies for unknown linear discrete-time systems from online data. The proposed approach views the data-dependent control synthesis problem as a time-varying semidefinite program (SDP) whose coefficients are recursively updated from online closed-loop measurements. Instead of repeatedly solving a full SDP as new data arrive, the policy is updated online through lightweight primal-dual iterations, each consisting of a linear equation solve and a projection onto the positive semidefinite cone. The framework applies to both direct and indirect data-driven formulations and covers a broad class of control objectives, including LQR, $H_\infty$ control, and safety-critical control. To characterize the coupling between online optimization and closed-loop data generation, we introduce two data-dependent quantities: the Sim-to-Real Gap, which measures the mismatch between noisy and noiseless data-induced SDPs, and the Difference-of-Signal, which measures the temporal variation of the SDP coefficients. Under persistency of excitation, suitable SDP regularity conditions, and sufficiently slow data variation, we establish a local linear tracking result up to residual terms governed by the latter two quantities. A global ergodic convergence bound is also derived for arbitrary initialization. Numerical examples on LQR, $H_\infty$ control, and safe exploration demonstrate that the proposed method can efficiently improve control performance from online data while accommodating SDP constraints beyond the well-explored LQR policy-gradient formulations.

View free PDFSource page

Related papers

arxivmath.OCcs.LGeess.SY2026-07-14

Learning-enabled Acceleration of Scenario-based Model Predictive Control

Trinh Tran, Binh Nguyen, Truong X. Nghiem

Scenario-based model predictive control (SBMPC) is a variant of model predictive control (MPC) that explicitly accounts for uncertainty by optimizing control actions over multiple predicted scenarios. However, its computational complexity increases rapidly with the number of scen…

View free PDFSource page
arxivquant-pheess.SYmath.OC2026-07-03

Nested-Loop Trajectory-Informed Variational Quantum Solver for Interior-Point OPF

Farshad Amani, Amin Kargarian

Optimal power flow (OPF) solved by an interior-point method (IPM) requires repeatedly solving Newton linear systems. When variational quantum linear solvers (VQLS) are used, each IPM iteration involves an additional nested inner variational optimization loop, which can significan…

View free PDFSource page
arxiveess.SYmath.OC2026-07-17

Gaussian behaviors and stochastic data-driven control

András Sasfi, Alberto Padoan, Ivan Markovsky, Florian Dörfler

We propose a stochastic behavioral modeling framework, termed Gaussian behaviors, which augments a deterministic linear time-invariant (LTI) behavior with a Gaussian noise component. We show that this notion is a tractable subclass of stochastic behaviors and encompasses classica…

View free PDFSource page
arxiveess.SYcs.LGmath.OC2026-07-17

Pick-to-Learn Calibration of an MPC Policy for an Origin-to-Destination Flight Problem

Marco C. Campi, Simone Garatti

This paper illustrates the Pick-to-Learn methodology applied to the calibration of a Model Predictive Control policy. While developed around a specific example, the presentation is meant to highlight a methodology of broad applicability. The example concerns an aircraft traveling…

View free PDFSource page
arxivmath.OCeess.SY2026-07-17

Smoothed Two-Stage Decomposition Algorithm for Solving Large-Scale Transmission and Distribution AC-OPF Problems

Juan Ospina, Manuel Garcia, Xinyi Luo, Andreas Wachter, David M. Fobes, Russell Bent

The integration of distributed energy resources (DERs) into the power grid has introduced new challenges to AC optimal power flow (AC-OPF) problems. Traditional OPF optimize consider transmission systems, treating distribution networks as static loads. However, the growing presen…

View free PDFSource page
arxiveess.SYmath.OC2026-07-21

Model-Agnostic Meta Learning for Differentiable MPC

Salma Elfeki, Riccardo Zuliani, Niklas Schmid, Efe C. Balta, John Lygeros

Applying policy optimization to Model Predictive Control (MPC) yields high-performance and reliable controllers. However, the resulting controllers often overfit their training conditions and suffer significant performance degradation in unseen tasks. We propose a novel framework…

View free PDFSource page