Search arXivSearch

arXiv · 2004.00241

Regret Bounds for LQ Adaptive Control Under Database Attacks (Extended Version)

Abstract

This paper is concerned with understanding and countering the effects of database attacks on a learning-based linear quadratic adaptive controller. This attack targets neither sensors nor actuators, but just poisons the learning algorithm and parameter estimator that is part of the regulation scheme. We focus on the adaptive optimal control algorithm introduced by Abbasi-Yadkori and Szepesvari and provide regret analysis in the presence of attacks as well as modifications that mitigate their effects. A core step of this algorithm is the self-regularized on-line least squares estimation, which determines a tight confidence set around the true parameters of the system with high probability. In the absence of malicious data injection, this set provides an appropriate estimate of parameters for the aim of control design. However, in the presence of attack, this confidence set is not reliable anymore. Hence, we first tackle the question of how to adjust the confidence set so that it can compensate for the effect of the poisonous data. Then, we quantify the deleterious effect of this type of attack on the optimality of control policy by bounding regret of the closed-loop system under attack.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jafar Abbaszadeh Chekan, Cedric Langbort. 2022-11-05. Regret Bounds for LQ Adaptive Control Under Database Attacks (Extended Version). https://arxiv.org/abs/2004.00241

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Cooperative Multi-Agent Assignment over Stochastic Graphs via Constrained Reinforcement Learning

Constrained multi-agent reinforcement learning offers the framework to design scalable and almost surely feasible solutions for teams of agents operating in dynamic environments to carry out conflicting tasks. We address the challenges of multi-agent coordination through an unconventional formulation in which the dual variables are not driven to convergence but are free to cycle, enabling agents to adapt their policies dynamically based on real-time constraint satisfaction levels. The coordination relies on a light single-bit communication protocol over a network with stochastic connectivity. Using this gossiped information, agents update local estimates of the dual variables. Furthermore, we modify the local dual dynamics by introducing a contraction factor, which lets us use finite communication buffers and keep the estimation error bounded. Under this model, we provide theoretical guarantees of almost sure feasibility and corroborate them with numerical experiments in which a team of robots successfully patrols multiple regions, communicating under a time-varying ad-hoc network.

eess.SY

Which Top Energy-Intensive Manufacturing Countries Can Compete in a Renewable Energy Future?

In a world increasingly powered by renewables and aiming for greenhouse gas-neutral industrial production, the future competitiveness of todays top manufacturing countries is questioned. This study applies detailed energy system modeling to quantify the Renewable Pull, an incentive for industry relocation exerted by countries with favorable renewable conditions. Results reveal that the Renewable Pull is not a cross-industrial phenomenon but strongly depends on the relationship between energy costs and transport costs. The intensity of the Renewable Pull varies, with China, India, and Japan facing a significantly stronger effect than Germany and the United States. Incorporating national capital cost assumptions proves critical, reducing Germanys Renewable Pull by a factor of six and positioning it as the second least affected top manufacturing country after Saudi Arabia. Using Germany as a case study, the analysis moreover illustrates that targeted import strategies, especially within the EU, can nearly eliminate the Renewable Pull, offering policymakers clear options for risk mitigation.

eess.SY

Certifying Frequency Stability for Systems with Line Dynamics and Heterogeneous Bus Dynamics

This work presents a framework for certifying small-signal frequency stability of a power system with line dynamics and heterogeneous bus dynamics. This framework can certify the stability of systems which include synchronous generators, synchronous condensers, and converter-interfaced resources with a wide range of controls. Moreover, it can do so without detailed or precise knowledge of the network topology. With this framework, we also provide a detailed analysis of how proportional-derivative (PD) droop can improve the stability margin of the frequency response. The stability certificates presented in this work, which extend prior results by incorporating line dynamics, provide insight into how the control parameters for different units in the system impact the overall frequency stability. While damper windings have long been understood to improve the frequency synchronization between machines, the dynamics of the damper windings are complex, making them difficult to analyze. To address this gap, this paper derives a novel reduced-order model of the damper windings in the form of a derivative droop term. Moreover, we show that derivative droop terms used in grid-forming (GFM) control can be understood as a form of damper winding emulation. Our analytical stability conditions highlight the importance of damper windings (or their emulation) in facilitating frequency synchronization and suppressing unstable interactions between GFM converters. These results are validated with electromagnetic-transient (EMT) simulation.

eess.SY