Skip to content
Open access

RL-SDNTE: Reinforcement Learning-Driven Traffic Engineering in SDN for QoE Optimization in Video Streaming

Saurabh Suman Roopali Lolag Sanjay Sange Sonali Padalkar
Aug 2026 · International journal of computer information systems and industrial management applications · 0 citations

TL;DR

RL-SDNTE is presented, a Reinforcement Learning-based TE framework built directly into an SDN controller that targets end-user Quality of Experience (QoE) as its primary objective and scales to topologies beyond 100 nodes without exceeding operationally acceptable convergence times.

Abstract

Video streaming now accounts for over 80% of global Internet bandwidth, yet most SDN traffic engineering (TE) solutions still optimize for throughput and link utilization rather than what users actually experience. Poor startup times, frequent re-buffering, and unstable bit-rate remain common even on well-managed networks -- largely because the control plane has no visibility into application-layer quality. We present RL-SDNTE, a Reinforcement Learning-based TE framework built directly into an SDN controller that targets end-user Quality of Experience (QoE) as its primary objective. Rather than relying on a single proxy metric, RL-SDNTE feeds four perceptual indicators -- startup latency, re-buffering ratio, mean video quality, and bit-rate oscillation -- into a unified reward function that drives routing decisions. A Deep Q-Network (DQN) agent uses the controller’s global network view together with real-time client feedback to continuously adjust path selection. Testing on a Mini-net emulation platform and a physical 12-node SDN test-bed showed gains of up to 34% in composite QoE, 28% fewer re-buffering events, 22% lower startup latency, and 17% less quality oscillation compared to ECMP, OSPF, DEFO, and heuristic QoE- aware baselines [5]-[7],[15]. The system also scales to topologies beyond 100 nodes without exceeding operationally acceptable convergence times, making it viable for real-world SDN deployments.

Read PDF

Similar papers

2026

PPO-MS: Confidence-Aware and Collaborative Traffic Management for Multimedia Streaming in SDN

The rapid growth of multimedia streaming poses critical challenges, including bursty traffic and congestion, leading to playback delays. The existing separate prediction and control mechanisms for multimedia traffic scheduling, which are based on software-defined networks (SDN), are unable to proactively manage bursty...

Jia-Wei Wu, Yibo Wang, Zelin Zhu · 0 citations
Open access Jul 2026

STQ-Scheduler: A Secure and Throughput-Aware Deep Reinforcement Learning Framework for QoE-Driven Resource Scheduling in Distributed Video Streaming Systems

STQ-Scheduler is proposed, a secure and throughput-aware deep reinforcement learning framework that integrates high-throughput data processing, Transformer-based QoE prediction, and Proximal Policy Optimization-based resource scheduling to ensure data quality and prevent data processing from becoming a bottleneck in di...

Yi-Chun Chang, Min-Wei Jiang · 0 citations
Open access Jul 2026

Robust Offline Multi-Agent Reinforcement Learning for Latency-Aware SDN Path Control in 6G-Oriented Network Softwarization

Future sixth-generation (6G)-oriented networks require programmable control that can adapt routing to latency and congestion without unsafe online exploration. This study evaluates offline multi-agent deep deterministic policy gradient (MADDPG) with behavior-adjusted training rewards for latency-aware path control in s...

A. Kyzyrkanov, Y. Nurakhov, Zhenis Otarbay et al. · 0 citations
Open access Aug 2026

Prioritized Experience Replay-Based Deep Deterministic Policy Gradient for Reliable Path Selection in SDN-IoT Networks

: Routing optimization is becoming prominent in Software-Defined Networks (SDN) due to the exponential growth of network traffic demands and the requirement for Quality of Service (QoS). However, reliable routing that satisfies the QoS requirements, such as end-to-end delay, packet loss, and bandwidth, remains a diffic...

Gaurav Kumar, G. Girisha, N. Shamanth · 0 citations
Review Open access 2026

Comprehensive Review of Optimization Techniques for User-Centric Distributed Network Slicing in 5G Networks

A QoE-aware framework for Multi-Access Edge Computing-enabled Open Radio Access Network (O-RAN) architectures, combining a graph attention network (GAT) encoder, distributed multi-agent DRL, and privacy-preserving FL, while transitioning control from Quality of Service (QoS) to QoE metrics is proposed.

Manoj Prasad Kunasegran, Wai Leong Pang, S. K. Phang · 0 citations
Conference Jul 2026

TS-D3QS: A Traffic-State-Aware Dueling Double-DQN Scheduler for Adaptive Network Queue Control

Adaptive queue management must balance throughput, delay, packet loss, and fairness under changing traffic and resource conditions. This paper proposes TS-D3QS, a traffic-state-aware queue scheduler that formulates multi-queue resource allocation as a discrete reinforcement-learning problem. The scheduler observes norm...

Hao-Yan Wang, Q. Guan, Dapeng Yan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.