SearcharxivSearch

arXiv subjects

Silvia Cianchi

Publications and source records attributed to Silvia Cianchi.

2 recordsLinked to original sources

Learning-Based Stackelberg Equilibrium Seeking with Application to Demand-Side Energy Management

Demand-side management (DSM) enables distribution system operators (DSOs) to steer electricity consumption through dynamic price signals or incentive mechanisms, thereby leveraging end-users' flexibility potential for delivering grid services. The resulting hierarchical interaction between the DSO and the end-users can be formulated as a Stackelberg game, where the operator dynamically sets the prices and the end-users optimally respond to them. Efficiently designing these price signals is challenging, as the users' response models are unknown or difficult to estimate. In this paper, we propose a learning-based zeroth-order algorithm for incentive design, in which the iterative update of the incentive signals is efficiently assisted by a data-driven online estimation of the users' responses. The proposed method is then proven to converge to an equilibrium tariff while allowing the DSO to estimate the decision-making problems at the user level. Moreover, the method preserves users' privacy, as the update rule of the DSO is solely based on observations of communicated end-user actions. Numerical simulations employing real-world data illustrate the efficient convergence of our learning-based proposed method, while significantly reducing the number of required interactions between the DSO and the end-users with respect to the state-of-the-art approach.

math.OC

Induced Stackelberg Equilibrium Seeking via Iterative Tikhonov Regularization

Existing methods for learning Stackelberg equilibria typically assume that the followers' (variational, generalized) Nash equilibrium is unique. However, in the presence of multiple equilibria, without a selection convention, the problem may become ill-posed, thus leading standard algorithms to potentially fail to converge. This paper addresses this issue by introducing an optimal selection at the lower-level game, hereby defining a Stackelberg game with induced equilibrium selection. To this end, we enable the leader to augment the followers' game with an additional vanishing term that acts as an incentive. We then propose a follower-agnostic zeroth-order method, whereby the leader converges to a solution of the resulting problem by iteratively probing the followers and jointly updating its decision variable and the incentive term.

math.OC