Skip to content
Open access

Reliability-Aware Local-Grid-Based Multipath Routing with Q-Learning Adaptation for Wireless Sensor Networks with a Mobile Sink

Aug 2026 · Applied Sciences · 0 citations · 25 references

TL;DR

QL-LGMPRP is proposed, a reliability-aware local-grid-based multipath routing protocol that combines a sink-centered local grid, two-path delivery, link-quality-aware forwarding, and lightweight tabular Q-learning for waypoint adaptation.

Abstract

Multipath routing in wireless sensor networks (WSNs) improves reliability by providing alternative forwarding paths when a route fails. However, mobile sinks make path maintenance difficult because sink movement can invalidate previously constructed source-to-sink routes. Existing protocols typically depend on either global path reconstruction, which increases control overhead, or footprint-chaining, which accumulates detours through previous sink positions and may weaken path independence. To address this problem, this paper proposes QL-LGMPRP, a reliability-aware local-grid-based multipath routing protocol that combines a sink-centered local grid, two-path delivery, link-quality-aware forwarding, and lightweight tabular Q-learning for waypoint adaptation. Mobility-related route changes are confined to the sink-centered grid, whereas a compact tabular Q-learning policy adjusts the primary-path direction using grid, link-quality, and energy-related state variables. The sink constructs a local grid around its current position, with cells sized to keep in-grid forwarding locally bounded. When an event occurs, the source computes an entry point on the grid perimeter and constructs two greedy paths: a primary path through a Q-learning-selected waypoint near the grid boundary and a backup path toward the current sink position. The Q-learning agent uses a compact tabular state representation that includes the boundary-cell index, residual-energy level, sink-grid position, and local link-quality information, and learns waypoint offsets using a reward that combines delivery success, transmission energy, and delay. This design confines routing adaptation to the sink-centered grid while allowing the waypoint policy to respond to heterogeneous link conditions. Simulation results under different sink speeds and interference conditions show that QL-LGMPRP maintains high delivery reliability while reducing detour-related forwarding costs relative to footprint-chaining and showing lower weak-link exposure than the geometric-forwarding comparison schemes.

Read PDF

Similar papers

#reinforcement learning Open access Sep 2026

A Q-learning-based mobility-aware tree routing protocol in wireless body area networks

This study proposes a Q-learning-based mobility-aware tree routing protocol (QMTR) for WBANs integrated with vehicular ad hoc networks (VANETs) to support reliable transmission of passengers’ health data to remote medical centers.

Mehdi Hosseinzadeh, Jawad Tanveer, Amir Masoud Rahmani et al. · 0 citations
Open access Aug 2026

TLQ-Geo: a two-level q-learning-based geographic routing protocol for flying ad hoc networks

A two-level Q-learning-based geographic routing protocol called TLQ-Geo for FANETs, which significantly reduces convergence time and computational overhead and integrates hierarchical decision-making with adaptive reinforcement learning.

Mehdi Hosseinzadeh, Jawad Tanveer, Amir Masoud Rahmani et al. · 0 citations
Open access Sep 2026

QoS-Aware Multi-Hop Routing Framework for Reliable Wireless Communication in Dynamic Ad Hoc Networks

Mobile Ad Hoc Networks (MANETs) enable infrastructure-free wireless communication by allowing mobile nodes to cooperate as routers over multiple hops. Their flexibility is valuable in emergency response, tactical communication, temporary field networks, disaster recovery, and mobile sensing, but reliable Quality of Ser...

Dr. Malothu Amru, Dr. Bitla Prabhakar, Somala Rama Kishore T. Satyanarayana · 0 citations
Open access 2026

Feasibility-Aware Reinforcement Learning for Reliable Hop-Constrained Routing in Wireless Sensor Networks

The experiments show that feasibility-aware learning can approach deterministic baseline reliability while retaining learned forwarding capability under hop constraints, and confirm that action masking is the dominant mechanism for maintaining feasible routing decisions, whereas trust mainly provides reliability-aware...

Adeel Iqbal, Muhammad Faisal Siddiqui · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.