Aug 2026· Applied Sciences· 0 citations· 25 references
TL;DR
QL-LGMPRP is proposed, a reliability-aware local-grid-based multipath routing protocol that combines a sink-centered local grid, two-path delivery, link-quality-aware forwarding, and lightweight tabular Q-learning for waypoint adaptation.
Abstract
Multipath routing in wireless sensor networks (WSNs) improves reliability by providing alternative forwarding paths when a route fails. However, mobile sinks make path maintenance difficult because sink movement can invalidate previously constructed source-to-sink routes. Existing protocols typically depend on either global path reconstruction, which increases control overhead, or footprint-chaining, which accumulates detours through previous sink positions and may weaken path independence. To address this problem, this paper proposes QL-LGMPRP, a reliability-aware local-grid-based multipath routing protocol that combines a sink-centered local grid, two-path delivery, link-quality-aware forwarding, and lightweight tabular Q-learning for waypoint adaptation. Mobility-related route changes are confined to the sink-centered grid, whereas a compact tabular Q-learning policy adjusts the primary-path direction using grid, link-quality, and energy-related state variables. The sink constructs a local grid around its current position, with cells sized to keep in-grid forwarding locally bounded. When an event occurs, the source computes an entry point on the grid perimeter and constructs two greedy paths: a primary path through a Q-learning-selected waypoint near the grid boundary and a backup path toward the current sink position. The Q-learning agent uses a compact tabular state representation that includes the boundary-cell index, residual-energy level, sink-grid position, and local link-quality information, and learns waypoint offsets using a reward that combines delivery success, transmission energy, and delay. This design confines routing adaptation to the sink-centered grid while allowing the waypoint policy to respond to heterogeneous link conditions. Simulation results under different sink speeds and interference conditions show that QL-LGMPRP maintains high delivery reliability while reducing detour-related forwarding costs relative to footprint-chaining and showing lower weak-link exposure than the geometric-forwarding comparison schemes.
This study proposes a Q-learning-based mobility-aware tree routing protocol (QMTR) for WBANs integrated with vehicular ad hoc networks (VANETs) to support reliable transmission of passengers’ health data to remote medical centers.
Mehdi Hosseinzadeh, Jawad Tanveer, Amir Masoud Rahmani et al.· Journal of King Saud Univers...· 0 citations
A two-level Q-learning-based geographic routing protocol called TLQ-Geo for FANETs, which significantly reduces convergence time and computational overhead and integrates hierarchical decision-making with adaptive reinforcement learning.
Mehdi Hosseinzadeh, Jawad Tanveer, Amir Masoud Rahmani et al.· Journal of King Saud Univers...· 0 citations
Mobile Ad Hoc Networks (MANETs) enable infrastructure-free wireless communication by allowing mobile nodes to cooperate as routers over multiple hops. Their flexibility is valuable in emergency response, tactical communication, temporary field networks, disaster recovery, and mobile sensing, but reliable Quality of Ser...
Dr. Malothu Amru, Dr. Bitla Prabhakar, Somala Rama Kishore T. Satyanarayana· International Journal of Adv...· 0 citations
The experiments show that feasibility-aware learning can approach deterministic baseline reliability while retaining learned forwarding capability under hop constraints, and confirm that action masking is the dominant mechanism for maintaining feasible routing decisions, whereas trust mainly provides reliability-aware...
Adeel Iqbal, Muhammad Faisal Siddiqui· Computers, Materials & C...· 0 citations