Skip to content
Open access

Lightweight reinforcement learning for leader following in autonomous vehicle convoys

Aug 2026 · Scientific Reports · 0 citations

TL;DR

This work presents an innovative strategy for autonomous convoy navigation that rethinks traditional approaches and designates a single lead vehicle to perform the heavy lifting of global path planning and obstacle mapping and refers to this leader-follower framework as the Pied Piper Policy (PPP), reflecting its divide-and-conquer approach to convoy autonomy.

Abstract

This work, presents an innovative strategy for autonomous convoy navigation that rethinks traditional approaches. Instead of equipping each vehicle with its own computationally intensive navigation system, our framework designates a single lead vehicle to perform the heavy lifting of global path planning and obstacle mapping. Follower vehicles adopt a lightweight, reinforcement learning (RL)–driven policy that allows for rapid local adaptations. Extensive validation through simulations in both PyBullet and Gazebo environments demonstrates that our decentralized scheme markedly reduces overall computational overhead and energy consumption while enhancing the synchronization and responsiveness of the convoy. We refer to this leader-follower framework throughout this paper as the Pied Piper Policy (PPP), reflecting its divide-and-conquer approach to convoy autonomy.

Read PDF

Similar papers

Open access Aug 2026

Deep Reinforcement Learning for Communication-Free Distributed Control of Autonomous Vehicles in Unstructured Intersection.

A novel deep reinforcement learning framework that enables safe and efficient navigation in such communication-free, signal-free, and lane-free intersection environments while meeting stringent safety requirements for practical deployment is proposed.

Ruihao Zeng, M. Ramezani · 0 citations
Open access Sep 2026

A platform for the navigation of Unmanned Aerial Vehiclesbased on Reinforcement Learning

This article introduces a platform for developing use cases for the automated control of unmanned aerial vehicles (UAVs), utilising the AirSim simulator. This platform allows for the generation of realistic flight scenarios involving multiple UAVs.The proposed platform facilitates the construction of use cases for the...

Daniel Álvarez Novoa, Ángel-Grover Pérez-Muñoz, Alejandro Alonso et al. · 0 citations
Open access Aug 2026

SkyAgent: A lightweight LLM-driven reinforcement learning framework for adaptive cooperative path planning of two UAVs

This work provides a feasible technical pathway and reproducible evaluation benchmark for the collaborative deployment of lightweight LLM planner, sub-goal guidance, sensor observations, cooperative reward, and reward shaping components and quantifies the indispensability of the LLM planner.

Yuting Cao, Zheng Zhao, Jiekai Wu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Decentralized Safe Multi-Agent Reinforcement Learning via Predictive Shielding

Environments are increasingly populated by multiple robots performing independent tasks with limited prior knowledge of each other. Deploying such multi-agent systems presents significant challenges. Specifically, shifts in deployment states compared to training data can lead to poor policy performance and compromised...

Y. El Yamani, Hanna Krasowski, Elena Vanneaux · 0 citations
Conference Sep 2026

Intelligent autonomous navigation for UAVs: a strategic hierarchical path planning framework

To address the challenge that single algorithms struggle to balance global exploration and local obstacle avoidance, and are prone to falling into local optima in complex environments, this paper proposes a Strategic Hierarchical Path Planning (SHPP) framework. This framework decouples the 3D navigation task into three...

Fei Wang, Jun-Yong Shi, Zhao-Kun Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.