Aug 2026· International Journal of Wireless and Microwave Technologies· 0 citations
TL;DR
Simulations across various 5G IoT spectrum environments showed that F-DMRL performed faster adaptation, higher spectral efficiency, and lower interference probability compared to centralized meta-RL, federated DRL, and traditional decentralized RL baselines.
Abstract
Dynamic spectrum access (DSA) in 5G IoT setups with cognitive radio is characterized by rapid and decentralized decision-making processes in highly non-stationary wireless environments, limited communication needs, and restrictive bounds. In this work, we present F-DMRL, a federated, communication-efficient decentralized meta-reinforcement learning framework for allowing a massive number of IoT devices to meta-learn collectively about spectrum-access strategies in a decentralized way without centralized control and without an extensive amount of inter-agent communication. Our method incorporates lightweight federated meta-parameter aggregation with gradient sparsification and periodic communication, allowing devices to only compress the meta-updates during this process and then adapt locally for task-specificity. We have presented analytical speedup guarantees and upper bounds on communication cost under bounded environmental drift and shown that using the approach proposed here, F-DMRL preserves convergence properties while posing a large reduction in coordination overhead at the same time. Simulations across various 5G IoT spectrum environments showed that F-DMRL performed faster adaptation (up to 45% fewer episodes), higher spectral efficiency, and lower interference probability compared to centralized meta-RL, federated DRL, and traditional decentralized RL baselines. Simulation results averaged across 10 independent runs demonstrate improvements of 45% faster adaptation and 60–80% lower communication overhead relative to baseline methods, while maintaining stable convergence.
A packet-level transmission framework that captures buffer overflow, delay violations, and transmission errors, and uses the resulting packet delivery ratio (PDR) to represent partial-update reception through a packetized, Bernoulli-masked FL aggregation process is developed.
Examination of aggregation stability and feasibility-sensitive aggregation in FRL for Edge-IoT systems identifies a need for aggregation mechanisms that jointly account for update stability, update reliability, resource availability, and constraint feasibility.
Majid A. Aslan, A. Al-Shalabi, Ahmed S. Alhegami· مجلة جامعة صنعاء للعلوم التط...· 0 citations
AF-EdgeRL is proposed, a novel Byzantine-resilient Asynchronous Federated Reinforcement Learning framework tailored for distributed resource allocation and dynamic task offloading and establishes theoretical convergence guarantees under non-convex reinforcement learning objectives.
Daniel Merrow, Tember L. Nair, Lucas Farnandez· International Journal of App...· 0 citations
The proposed framework is validated by conducting simulation-based experiments on the benchmark datasets and synthetic autonomous workloads, where the novelty lies in the design of the system-level federated learning architecture, instead of the datasets themselves.
Jyotsnarani Tripathy, D. Rajalakshmi, A. N. Ramya Shree et al.· SN Computer Science· 0 citations
In strongly interference-coupled reuse-$1$ simulations, FedCritic-MIMO achieves the best performance-communication tradeoff among heuristic, independent-learning, centralized-training, and communication-ablation baselines.
Collaborative training in distributed semantic communication (DSC) networks typically relies on decentralized federated learning (DFL). However, pushing topology-agnostic aggregation into heterogeneous, multi-task environments creates a fundamental bottleneck: it drives negative transfer and over-consensus bias (OCB)....
Linqi Yin, Tie-Jun Lv, Wei-Cai Li et al.· IEEE Transactions on Communi...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.