Skip to main navigation Skip to search Skip to main content

Toward Scalable Multirobot Control: Fast Policy Learning in Distributed MPC

Xinglong Zhang*, Wei Pan, Cong Li, Xin Xu*, Xiangke Wang, Ronghua Zhang, Dewen Hu

*Corresponding author for this work

Research output: Contribution to journalArticleScientificpeer-review

27 Downloads (Pure)

Abstract

Distributed model predictive control (DMPC) is promising in achieving optimal cooperative control in multirobot systems (MRS). However, real-time DMPC implementation relies on numerical optimization tools to periodically calculate local control sequences online. This process is computationally demanding and lacks scalability for large-scale, nonlinear MRS. This article proposes a novel distributed learning-based predictive control framework for scalable multirobot control. Unlike conventional DMPC methods that calculate open-loop control sequences, our approach centers around a computationally fast and efficient distributed policy learning algorithm that generates explicit closed-loop DMPC policies for MRS without using numerical solvers. The policy learning is executed incrementally and forward in time in each prediction interval through an online distributed actor-critic implementation. The control policies are successively updated in a receding-horizon manner, enabling fast and efficient policy learning with the closed-loop stability guarantee. The learned control policies could be deployed online to MRS with varying robot scales, enhancing scalability and transferability for large-scale MRS. Furthermore, we extend our methodology to address the multirobot safe learning challenge through a force field-inspired policy learning approach. We validate our approach's effectiveness, scalability, and efficiency through extensive experiments on cooperative tasks of large-scale wheeled robots and multirotor drones. Our results demonstrate the rapid learning and deployment of DMPC policies for MRS with scales up to 10 000 units.

Original languageEnglish
Pages (from-to)1491-1512
Number of pages22
JournalIEEE Transactions on Robotics
Volume41
DOIs
Publication statusPublished - 2025

Bibliographical note

Green Open Access added to TU Delft Institutional Repository 'You share, we take care!' - Taverne project https://www.openaccess.nl/en/you-share-we-take-care
Otherwise as indicated in the copyright section: the publisher is the copyright holder of this work and the author uses the Dutch legislation to make this work public.

Keywords

  • Distributed MPC
  • multirobot systems
  • policy learning
  • safe learning
  • scalability

Fingerprint

Dive into the research topics of 'Toward Scalable Multirobot Control: Fast Policy Learning in Distributed MPC'. Together they form a unique fingerprint.

Cite this