Skip to content

A study on deep reinforcement learning algorithms for inter-satellite communication link scheduling in low-Earth orbit

Jul 2026 · Digital Signal and Computer Communications · Vol 14294, pp. 142940M - 142940M-7 · 0 citations · 11 references
Engineering

TL;DR

Experimental results demonstrate that this link scheduling model based on deep reinforcement learning exhibits superior performance in reducing transmission delay, enhancing system throughput, and improving load balancing, thereby validating its effectiveness and adaptability in complex inter-satellite link scheduling scenarios.

Abstract

Artificial intelligence-driven intelligent algorithms have demonstrated excellent adaptive optimization capabilities in complex network scheduling problems. To address the issues of dynamic topological changes and insufficient resource allocation efficiency in LEO satellite inter-satellite communication link scheduling, this study proposes a link scheduling model based on deep reinforcement learning. Building upon a dynamic time-varying network model, a Markov decision process is introduced to describe the scheduling process, and a deep neural network is employed to achieve a nonlinear mapping from states to actions. By integrating delay, throughput, and load balancing metrics through a multi-objective reward function, the scheduling strategy is optimized. Experimental results demonstrate that this method exhibits superior performance in reducing transmission delay, enhancing system throughput, and improving load balancing, thereby validating its effectiveness and adaptability in complex inter-satellite link scheduling scenarios.

View source

Similar papers

2026

Distributed Routing for LEO Satellite Networks: A Multi-Agent Deep Reinforcement Learning Approach With State Information Lag

Multi-agent deep reinforcement learning (MADRL) offers a promising solution for routing in low Earth orbit (LEO) satellite networks. However, large inter-satellite propagation delays lead to severe state information lag in agent interactions, giving rise to decision biases and degraded routing timeliness. To this end,...

Wei-Dan Liu, Tong Liu, Li-Xia Xiao et al. · 0 citations
Conference Aug 2026

Transformer-Enabled Constrained Deep Reinforcement Learning for Joint Satellite Selection and Power Control in LEO-GEO Coexistence Networks

In recent years, with the large-scale deployment of low Earth orbit (LEO) satellites, the scale of multi-user uplink transmission in satellite networks has been continuously expanding. During the process of user-satellite association and cochannel spectrum reuse, it not only triggers significant cross-user interference...

Ning Xu, Fan Zhou, Hao-Ge Jia et al. · 0 citations
Conference Open access 2026

Deep Reinforcement Learning Driven Spectrum-Energy Joint Optimization

Spectrum efficiency (SE) and energy efficiency (EE) are two fundamental problems in the wireless resource management that need to be jointly optimized. Deep reinforcement learning (DRL) allows near-optimal policy learning via continuous interaction with the environment, which is suitable in complex, dynamic, and high-d...

Jia-Yi Liu · 0 citations
Sep 2026

MAFRL: A Multi-Agent Flow-Balance Reinforcement Learning Framework for Resource Allocation in LEO Satellite Systems

Low Earth orbit (LEO) satellite systems are expected to deliver ubiquitous broadband connectivity, but their dynamic topology and limited on-board resources challenge beam hopping (BH) scheduling and inter-satellite load balancing. This paper proposes a multi-satellite cooperative BH algorithm based on multi-agent flow...

Sheng-Tong Xie, Cheng Wang, Gao-Feng Cui et al. · 0 citations
Conference Aug 2026

Deep Reinforcement Learning-Based Sensing Freshness Optimization in AoI-Constrained Air-Ground Communication Systems

In space-air-ground integrated emergency communication networks (SAGIECNs), unmanned aerial vehicles (UAVs) periodically upload sensing data to ground facilities, but limited battery capacity and wireless spectrum resources make it difficult to ensure both long operational lifetime and timely data transmission. This pa...

Bo-Yu Wu, Fang-Wei Ye, Yi-Ming Huang et al. · 0 citations
Conference Aug 2026

Prediction-Driven Task Scheduling in Satellite IoT: A Constrained Reinforcement Learning Approach

The incorporation of Mobile Edge Computing into satellite systems is a highly promising approach to enabling large-scale intelligent Internet of Things services in remote areas. However, the high-speed mobility of satellites and the extreme scarcity of on-board resources pose significant challenges, as traditional reac...

Hao-Yuan Deng, Ning-Ning Cui, Shi Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.