Multi-robot cooperative control has been extensively studied using model-based distributed control methods. However, such control methods rely on sensing and perception modules in a sequential pipeline design, and the separation of perception and controls may cause processing latencies and compounding errors that affect control performance. End-to-end learning overcomes this limitation by implementing direct learning from onboard sensing data, with control commands output to the robots. Challenges exist in end-to-end learning for multi-robot cooperative control, and previous results are not scalable. We propose in this article a novel decentralized cooperative control method for multi-robot formations using deep neural networks, in which inter-robot communication is modeled by a graph neural network (GNN). Our method takes LiDAR sensor data as input, and the control policy is learned from demonstrations that are provided by an expert controller for decentralized formation control. Although it is trained with a fixed number of robots, the learned control policy is scalable. Evaluation in a robot simulator demonstrates the triangular formation behavior of multi-robot teams of different sizes under the learned control policy.
more »
« less
This content will become publicly available on September 1, 2026
Stabilization via Distributed Control
We consider the linear plant with multiple pairs of input and output channels, and solve the problem of designing a stabilizing controller consisting of a set of control stations, one for each channel pair, connected with an arbitrarily fixed network topology and communication dynamics. It is shown that such stabilizing distributed controller exists if and only if the fixed modes of the plant, augmented with the network communication, are stable; this leverages and extends the classical fixed mode result for decentralized control. Moreover, a condition is given to reveal exactly what connections are needed between control stations to achieve stabilizability via a distributed control. Thus, the feasibility of optimal distributed control synthesis can readily be checked or enforced before applying computationally expensive optimization methods. Stabilizing distributed controllers, if feasible, can be designed using decentralized control theories.
more »
« less
- Award ID(s):
- 2113528
- PAR ID:
- 10667037
- Publisher / Repository:
- IEEE
- Date Published:
- Journal Name:
- IEEE Transactions on Automatic Control
- ISSN:
- 0018-9286
- Page Range / eLocation ID:
- 1 to 8
- Format(s):
- Medium: X
- Sponsoring Org:
- National Science Foundation
More Like this
-
-
Frequency restoration in power systems is conventionally performed by broadcasting a centralized signal to local controllers. As a result of the energy transition, technological advances, and the scientific interest in distributed control and optimization methods, a plethora of distributed frequency control strategies have been proposed recently that rely on communication amongst local controllers. In this paper, we propose a fully decentralized leaky integral controller for frequency restoration that is derived from a classic lag element. We study steady-state, asymptotic optimality, nominal stability, input-to-state stability, noise rejection, transient performance, and robustness properties of this controller in closed loop with a nonlinear and multivariable power system model. We demonstrate that the leaky integral controller can strike an acceptable trade-off between performance and robustness as well as between asymptotic disturbance rejection and transient convergence rate by tuning its DC gain and time constant. We compare our findings to conventional decentralized integral control and distributed- averaging-based integral control in theory and simulations.more » « less
-
We consider the problem of controlling a set of dynamically decoupled plants where the plants' subcontrollers communicate with each other according to a fixed and known network topology. We assume the communication to be instantaneous but there is a fixed processing delay associated with incoming transmissions. We provide explicit closed-form expressions for the optimal decentralized controller under these communication constraints and using standard LQG assumptions for the plants and cost function. Although this problem is convex, it is challenging due to the irrationality of continuous-time delays and the decentralized information-sharing pattern. We show that the optimal subcontrollers each have an observer-regulator architecture containing LTI and FIR blocks and we characterize the signals that subcontrollers should transmit to each other across the network.more » « less
-
Network-wide signal control optimization is of practical importance to shorten and stabilize travel time, improve productivity, enhance energy consumption efficiency, mitigate congestion, and reduce vehicle emissions. In this study, a deep learning-empowered distributed control strategy is developed to adaptively optimize network-wide traffic signal control coordination. To simplify the problem formulation and enhance its applicability, the entire traffic system is decomposed into multiple areas, and multilayer perceptron concepts are used to formulate traffic control system operations in each area. The distributed deep learning, velocity-based model predictive control (MPC) strategy is designed to optimize traffic signal coordination. Furthermore, a gain-scheduling control model is developed to linearize each learned nonlinear system around its most recent operating status, and then a distributed MPG controller is applied to the linearized systems. Simulation results demonstrate that the proposed control strategy can effectively reduce travel time by 15.1% compared with fixed-time control plans and by 8.0% compared with a decentralized control plan. This study is the first research effort to integrate the deep learning framework and multiagent MPG to optimize traffic control coordination. Moreover, a sufficient condition is theoretically formulated for the bounded-input, bounded-output stability of the closed-loop, large-scale traffic system based on the nonlinear small-gain theorem.more » « less
-
Control of networked systems, comprised of interacting agents, is often achieved through modeling the underlying interactions. Constructing accurate models of such interactions–in the meantime–can become prohibitive in applications. Data-driven control methods avoid such complications by directly synthesizing a controller from the observed data. In this paper, we propose an algorithm referred to as Data-driven Structured Policy Iteration (D2SPI), for synthesizing an efficient feedback mechanism that respects the sparsity pattern induced by the underlying interaction network. In particular, our algorithm uses temporary “auxiliary” communication links in order to enable the required information exchange on a (smaller) sub-network during the “learning phase”—links that will be removed subsequently for the final distributed feedback synthesis. We then proceed to show that the learned policy results in a stabilizing structured policy for the entire network. Our analysis is then followed by showing the stability and convergence of the proposed distributed policies throughout the learning phase, exploiting a construct referred to as the “Patterned monoid.” The performance of D2SPI is then demonstrated using representative simulation scenarios.more » « less
An official website of the United States government
