Skip to content
Open access

Learning Efficient Communication Protocols for Multi-Agent Reinforcement Learning

Nov 2025 · IEEE Transactions on Machine Learning in Communications and Networking · Vol 4, pp. 1335-1352 · 1 citation · 40 references
Computer Science

TL;DR

This work introduces a unified framework for learning multi-round communication protocols that are both effective and efficient and demonstrates that the learned communication protocols can significantly enhance communication efficiency and achieves better cooperation performance with improved success rates.

Abstract

Multi-Agent Systems (MAS) have emerged as a powerful paradigm for modeling complex interactions among autonomous entities in distributed environments. In Multi-Agent Reinforcement Learning (MARL), communication enables coordination but can lead to inefficient information exchange, since agents may generate redundant or non-essential messages. While prior work has focused on boosting task performance with information exchange, the existing research lacks a thorough investigation of both the appropriate definition and the optimization of communication protocols (communication topology and message). To fill this gap, we introduce a unified framework for learning multi-round communication protocols that are both effective and efficient. Within this framework, we propose three novel Communication Efficiency Metrics (CEMs) to guide and evaluate the learning process: the Information Entropy Efficiency Index (IEI) and Specialization Efficiency Index (SEI) for efficiency-augmented optimization, and the Topology Efficiency Index (TEI) for explicit evaluation. We integrate IEI and SEI as the adjusted loss functions to promote informative messaging and role specialization, while using TEI to quantify the trade-off between communication volume and task performance. Through comprehensive experiments, we demonstrate that our learned communication protocols can significantly enhance communication efficiency and achieves better cooperation performance with improved success rates.

Read PDF

Similar papers

#artificial intelligence Preprint Sep 2026

Robust and Efficient Communication for Multi-Agent Learning

Effective communication is a cornerstone of distributed intelligence in Multi-Agent Reinforcement Learning (MARL), yet ensuring that generated messages are both informative and robust to physical constraints remains a significant challenge. This paper introduces Multi-Agent Regularized Communication (MARC), a novel fra...

Rafael Pina, Varuna De Silva, Corentin Artaud · 0 citations
Sep 2026

Efficient Communication With Skill Neurons in Decentralized Multi-Agent Reinforcement Learning.

CSN is proposed, which enables efficient Communication with Skill Neurons in decentralized MARL by exchanging the essential components of learned knowledge at neuron level by communicating only a sparse subset of model parameters and doing so intermittently.

Jiahua Lan, Li Shen, Ruijun Liu et al. · 0 citations
Open access Aug 2026

Information Bottleneck for Communication-Efficient Multi-Agent Reinforcement Learning in UAV Swarms

IB-CEMARL is proposed, an information-bottleneck-guided, communication-efficient multi-agent reinforcement learning framework for UAV swarms that achieves superior cooperative performance, reduced message redundancy, and stronger robustness compared with representative communication-aware MARL baselines.

Zheng Yang, Guohao Li, Yali Xue · 0 citations
Conference Open access Sep 2026

Towards Streamlined Learning and Search for Multi-Agent Optimization

Focusing on multi-agent path finding as an exemplary problem, this paper proposes to simplify two popular approaches to MAPF, namely multi-agent reinforcement learning and adaptive search, to enable seamless combination and transferability of methods without substantial engineering effort.

Thomy Phan · 0 citations
Open access Aug 2026

AGTA: Topology-Aware Sequential Decision-Making in Multi-Agent Reinforcement Learning

Action Generation with Topology Awareness (AGTA), a topology-aware sequential decision-making framework in MARL that integrates inter-agent correlation modeling with topology-guided decision-order optimization, and outperforms the state-of-the-art counterparts.

Kun Hu, Shanghua Wen, Wen-Di Wu et al. · 0 citations
Jul 2026

Efficient Heterogeneous Exploration with Mutual Policy Divergence Maximization for Multiagent Reinforcement Learning.

This work introduces a novel MARL framework, Multi-Agent Divergence Policy Optimization (MADPO) with Mutual Policy Divergence Maximization (Mutual PDM), and proposes a new extension of CCS divergence for measuring policy divergence of more than two agents, the Generalized Conditional Cauchy-Schwarz (GCCS) divergence.

Haowen Dou, Lujuan Dang, Mingfei Lu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.