Open access
Aug 2026
MEMORY-AUGMENTED REINFORCEMENT LEARNING FOR UAV NAVIGATION USING PPO-LSTM
Experimental outcomes show that the PPO-LSTM described herein achieves smoother paths, more robust reward convergence, and a much lower rate of collision than regular PPO, and generalizes to new environments with movable obstacles.
M. Haddad, Dhayaa Khudher
· Kufa journal of Engineering · 0 citations