An integrated framework that combines LLM fine-tuning using federated adaptive local low-rank adaptation (FedALoRA) with forecasting-driven autoscaling policies is proposed, making it well-suited for federated-LLM Autoscaling in dynamic edge environments.
Abstract
The growing deployment of applications in real-world edge environments is driving an increasing demand for large language model (LLM) training in distributed systems. This demand creates significant challenges for resource management, particularly in resource-constrained edge cloud environments. Federated learning (FL) enables privacy-preserving LLM fine-tuning across multiple clients, but it also introduces dynamic computational and communication loads that can strain system resources. To address these challenges, we propose an integrated framework that combines LLM fine-tuning using federated adaptive local low-rank adaptation (FedALoRA) with forecasting-driven autoscaling policies. Differential privacy (DP) is incorporated to ensure secure aggregation, balancing privacy guarantees with system performance. Using the WikiText-2 dataset for FL-based LLM training and the Bitbrains dataset for CPU usage forecasting, we evaluate deep learning models (LSTM, GRU, BiLSTM, CNN, CNN-LSTM, Transformer). Results demonstrate that FedALoRA enables fast federated convergence, reducing perplexity from 199.3 to 85.2 without DP and from 211.6 to 92.2 with DP, incurring only a 7–13% privacy overhead. Among forecasting models, CNN-LSTM achieves the lowest MSE (0.25) and consistently requires the fewest pods (as low as 11 replicas), highlighting its superior accuracy and resource efficiency for proactive autoscaling. Comparative analysis demonstrates that proactive autoscaling consistently outperforms reactive autoscaling, with non-DP settings achieving slightly faster adaptation and DP settings offering stronger privacy guarantees. Overall, the proposed framework balances performance, adaptability, and privacy, making it well-suited for federated-LLM Autoscaling in dynamic edge environments.
Federated Learning (FL) enables collaborative model training across distributed devices. A major concern in FL is how to operate it smoothly in resource-constrained environments, where this process must perform under strict operational constraints, such as cost or energy budgets. An often-overlooked aspect is the effec...
Anna Lackinger, P. Frangoudis, Andrea Morichetta et al.· 2026 International Conferenc...· 0 citations
Energy-Aware Adaptive Quantization and Freezing (EA-AQF), a unified framework that co-optimizes communication and computation, is presented, a unified framework that co-optimizes communication and computation and maintains robust convergence in highly heterogeneous tasks.
A Federated Predictive Predictive Load Balancing (FPLB) framework is proposed to combine Long Short-Term Memory (LSTM) workload forecasting with federated learning, which does not require fog nodes to share their operational data.
Arti Sharma, R. Mahapatra, Vineet Sharma et al.· Scientific Reports· 0 citations
Federated learning trains shared models without centralizing local data, but repeated model exchange can become a bottleneck in bandwidth-, energy-, and latency-constrained edge systems. This paper presents FEDCODE, a communication-aware federated learning simulation framework for reproducible evaluation of update repr...
E. Guberović, Igor Čavrak· Italian National Conference...· 0 citations
The Federated Parallel Scaling (FPS) algorithm is proposed, which jointly trains multiple sampled subnetworks in parallel with self-distillation so that larger sampled subnetworks can supervise smaller ones during local updates.
Jia-Xin Zhang, Xing-Wei Wang, Bo Yi et al.· 0 citations
This work proposes a data-Quality-aware aggregation framework by introducing an Evolutionary-computation-inspired de-sign into Federated learning ( FedEvoQ), with a lightweight dual-branch architecture.
Le-Ming Wu, Yao-Chu Jin, Han Yu et al.· Proceedings of the Thirty-Fi...· 0 citations
Related blog posts
MIT News · Artificial Intelligence· news.mit.eduOct 7, 2026
Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.
Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.
MIT News · Artificial Intelligence· news.mit.eduOct 6, 2026