Skip to content
Open access

Decentralized MARL for SDN Ground Station Cluster Selection in LEO Constellations Under Stochastic Weather

2026 · IEEE Open Journal of the Communications Society · Vol 7, pp. 7770-7789 · 0 citations · 34 references
Computer Science

TL;DR

A hybrid LEO-terrestrial architecture that integrates Software-Defined Networking ground station clusters with repeater-assisted reception and a structured fallback mechanism is proposed and results highlight the effectiveness of the proposed framework in enabling adaptive and delay-efficient control in next-generation LEO satellite systems.

Abstract

Low Earth Orbit (LEO) satellite constellations enable global low-latency connectivity but face challenges due to weather-dependent link variability, orbital dynamics, and heterogeneous ground infrastructure. In this paper, we propose a hybrid LEO-terrestrial architecture that integrates Software-Defined Networking (SDN) ground station clusters with repeater-assisted reception and a structured fallback mechanism. We develop a stochastic model that captures weather-driven reliability, correlated repeater behavior, and multi-path reception, and formulate an optimization problem to minimize the expected communication delay. To address the intractability of this problem in large-scale, partially observable environments, we design a fully decentralized Multi-Agent Reinforcement Learning (MARL) framework based on Proximal Policy Optimization (PPO), where each satellite makes decisions using only local observations. The reward is aligned with the analytical delay objective, ensuring consistency between the model and learned policies. Simulation results across diverse scenarios demonstrate that the proposed approach reduces the mean delay by 40-60% and significantly decreases fallback usage compared to baseline methods. These results highlight the effectiveness of the proposed framework in enabling adaptive and delay-efficient control in next-generation LEO satellite systems.

Read PDF

Similar papers

2026

A Heterogeneous Multiagent Reinforcement Learning Approach for Robust Uplink Beamforming in Maritime Satellite Communications

Maritime satellite communications (SATCOMs) are expected to support high-capacity ship-to-satellite uplinks for remote maritime services beyond terrestrial coverage, with low-Earth-orbit (LEO) satellites providing wide-area connectivity. However, robust uplink beamforming in LEO maritime SATCOMs is challenging because...

Huayuan Wang, Bodong Shang, Meixia Tao · 0 citations
2026

Distributed Routing for LEO Satellite Networks: A Multi-Agent Deep Reinforcement Learning Approach With State Information Lag

Multi-agent deep reinforcement learning (MADRL) offers a promising solution for routing in low Earth orbit (LEO) satellite networks. However, large inter-satellite propagation delays lead to severe state information lag in agent interactions, giving rise to decision biases and degraded routing timeliness. To this end,...

Wei-Dan Liu, Tong Liu, Li-Xia Xiao et al. · 0 citations
Preprint Aug 2026

LEO-Aware DRL Meta-Scheduler for 5G Non-Terrestrial Network Slicing

The integration of Low Earth Orbit (LEO) Non-Terrestrial Networks (NTNs) into 5G and upcoming 6G architectures introduces various challenges, including severe propagation delays, ultra-high base station mobility, and channel non-stationarity, complicating radio resource management of heterogeneous network slices. In th...

Víctor Vilchez, T. P. C. de Andrade, Edward Hinojosa et al. · 0 citations
Open access 2026

ShellMean-MAPPO: A Conflict-Aware MARL Framework for Downlink Resource Allocation in Multi-Shell LEO Satellite Networks

This work proposes a structured multi-agent reinforcement learning (MARL) framework based on multi-agent proximal policy optimization (MAPPO), termed ShellMean-MAPPO, for downlink resource allocation with explicit conflict resolution, and demonstrates its advantages over representative MARL schemes in terms of scheduli...

Li Zhen, Qi-Hao Zhang, Qing-Zhi Meng et al. · 0 citations
Preprint Aug 2026

Multi-Agent Reinforcement Learning for Joint Handover Management and Power Allocation in Multi-Orbit Satellite Networks

The proposed multi-agent reinforcement learning policy attains slightly higher throughput with fewer handovers by offloading a fraction of the users to the MEO and GEO layers, an emergent multi-orbit behavior that drives its favorable throughput and handover trade-off.

Yassine Afif, Ashutosh Balakrishnan, Philippe Martins et al. · 0 citations
Open access 2026

Reliable Low-Latency Task Offloading and Resource Allocation Method for Space-Air-Ground Integrated Networks

: Space-Air-Ground Integrated Networks (SAGIN) provide a multi-layered, wide-coverage computing infrastructure for distributed urban sensing systems. However, their heterogeneity and dynamics pose unprecedented challenges for task offloading and resource allocation. Existing methods struggle to simultaneously address t...

Fei-Yan Bu, Zheng Wang, Yong Pan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.