Skip to content
Open access

RoboSwarmCoordAI enables scalable swarm robotics coordination using multi agent reinforcement learning

Aug 2026 · Discover Computing · Vol 29 · 0 citations · 44 references

TL;DR

RoboSwarmCoordAI is a promising simulation-validated framework for adaptive swarm coordination, and future work will further validate RoboSwarmCoordAI on larger swarms and physical robotic platforms.

Abstract

The field of swarm robotics has become more and more popular as a decentralised solution to problems of coordination among multiple independent agents. The recent developments in multi-agent reinforcement learning (MARL) have made it possible for agents to learn cooperative behaviours when operating in a dynamic environment and even outperform the traditional rule-based or heuristic coordination strategies. However, coordination in practical MARL-based swarms is still difficult, as many approaches are not scalable, have high communication cost, unstable coordination with high swarm density, and lack integration of efficiency, robustness, and adaptability. This paper introduces a multi-agent reinforcement learning framework for scalable and communication-efficient swarm coordination called RoboSwarmCoordAI, which surpasses the limitations of the above approaches. The proposed framework uses three major components: a state-encoding module that is aware of the coordination requirements, an adaptive neighbourhood-filtering module to avoid redundant inter-agent communication and a hybrid reward function that weights local robot goals with respect to the global swarm performance. RoboSwarmCoordAI uses a centralised training and decentralised execution approach where agents can leverage global information for training, but local information for execution. The framework was tested in simulation in cooperative exploration, distributed target search and task allocation scenarios. When evaluated within the range of simulations tested, RoboSwarmCoordAI outperformed baseline methods with a task success rate of 95.8%, 2.7 collisions per episode, and an efficiency score of 91.6. The analyses of scalability and communication efficiency also demonstrate the stable coordination performance up to 50 robots, at a lower communications cost. As a result, it was found that RoboSwarmCoordAI is a promising simulation-validated framework for adaptive swarm coordination, and future work will further validate RoboSwarmCoordAI on larger swarms and physical robotic platforms.

Read PDF

Similar papers

Sep 2026

Efficient Communication With Skill Neurons in Decentralized Multi-Agent Reinforcement Learning.

CSN is proposed, which enables efficient Communication with Skill Neurons in decentralized MARL by exchanging the essential components of learned knowledge at neuron level by communicating only a sparse subset of model parameters and doing so intermittently.

Jiahua Lan, Li Shen, Ruijun Liu et al. · 0 citations
Conference Open access Sep 2026

Towards Streamlined Learning and Search for Multi-Agent Optimization

Focusing on multi-agent path finding as an exemplary problem, this paper proposes to simplify two popular approaches to MAPF, namely multi-agent reinforcement learning and adaptive search, to enable seamless combination and transferability of methods without substantial engineering effort.

Thomy Phan · 0 citations
Review Open access Aug 2026

Decentralized Coordination Architectures for Intelligent Agent Swarms

The discussion treats distributed consensus, event-triggered communication, resilient control, fault-tolerant design, and cognition-inspired adaptation as parts of one architecture problem.

Tianwen Ge · 0 citations
Conference Open access 2026

Reinforcement Learning for Quadrupedal Robot Control: Taxonomy, Sim-to-Real, Robustness, and Emerging Trends

. Quadrupedal robots exhibit strong mobility in complex environments where wheeled platforms often perform poorly, but their control remains difficult. In recent years, reinforcement learning (RL) has received growing attention in quadrupedal locomotion, as it supports direct policy optimization without relying entirel...

Chen Chang · 0 citations
#graph neural networks Open access Sep 2026

Graph Neural Network‐Based Reinforcement Learning for Decentralized Multi‐Robot Manipulation

In tightly cooperative manipulation tasks, robotic manipulators must follow collision‐free and coordinated trajectories. Existing multiagent learning frameworks often rely on centralized planners that provide strong coordination but fail to scale with larger teams. Alternatively, decentralized approaches offer better s...

Tong Chen, Bo Fu, Dawn M. Tilbury et al. · 0 citations
Open access Aug 2026

Modeling Dynamic Obstacle Avoidance Strategy of Drone Swarms Combined with Multi-Agent Reinforcement Learning

The proposed framework demonstrates robust scalability and real-time coordination capability for dynamic environments, while providing a reliable decision-making paradigm for intelligent multi-agent systems operating in communication-intensive and electromagnetically complex application scenarios.

X.-H. Fang, K. Chen, Cheng-Hao Ren et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.