Skip to content
Open access

A Rust-Implemented Dueling Double DQN for Fixed-Wing Autonomous Flight

Aug 2026 · Journal of Advances in Mathematics and Computer Science · Vol 41, pp. 163-178 · 0 citations

TL;DR

These findings support the feasibility of native Rust-based D3QN training for real-time fixed-wing simulation, while the observed inter-seed variability indicates that reward shaping and convergence robustness require further evaluation.

Abstract

Deep reinforcement learning for autonomous unmanned aerial vehicle control has largely been demonstrated with multirotor platforms and high-level machine-learning frameworks. This study presents a Dueling Double Deep Q-Network (D3QN) training pipeline implemented in Rust without an external machine-learning library and integrated with Godot 4 through GDExtension for fixed wing flight control. The controller addresses fixed-wing requirements, including airspeed maintenance, lift management, throttle regulation, stall avoidance, and coordinated turning. The network combines a duelling architecture, double Q-learning, prioritised experience replay, and three-step returns in a 512→256 hidden-layer configuration containing 142,088 parameters for the 16-dimensional input case. The agent selects among seven discrete actions and supports both 12-dimensional and 16-dimensional observation spaces through a cross-dimensional weight-transfer procedure. Training was conducted for 200 episodes using four random seeds. Across seeds, the mean best episodic reward was 4825±40, while the coefficient of variation for best reward was 0.8%. In the final 30 episodes, no crashes were recorded, although completion rates varied substantially between seeds. Airspeed remained within ±8 m/s of the 50 m/s target. Batch-64 gradient updates required less than 1 ms, representing an approximately 35-fold reduction in latency relative to the preceding GDScript implementation, and the reported runtime memory footprint remained below 50 MB. These findings support the feasibility of native Rust-based D3QN training for real-time fixed-wing simulation, while the observed inter-seed variability indicates that reward shaping and convergence robustness require further evaluation.

Read PDF

Similar papers

Open access Sep 2026

A Reinforcement-Learning-Based Hybrid SAC–LQR Framework for UR3e Multi-Waypoint Motion: System Design and Simulation-Based Evaluation

Artificial intelligence is increasingly being investigated for robot motion generation, while conventional methods remain effective for deterministic waypoint tasks. This study evaluates reinforcement learning as a state-conditioned joint-reference-generation layer at runtime, not as a replacement for classical control...

Ahmed Iqdymat, I. Stamatescu, G. Stamatescu · 0 citations
Open access Aug 2026

Explainable Reinforcement Learning Framework for Autonomous Windshear Escape with Policy Distillation

This study proposes an explainable, data-driven framework integrating active-reward proximal policy optimization (AR-PPO), which successfully distills black-box AI strategies into verifiable, physics-informed standard operating procedures (SOPs), providing a highly transparent and robust solution for autonomous windshe...

Yi-Tan Wang, Yang-Yang Zhang, Zhen-Xing Gao · 0 citations
Preprint Sep 2026

Battery-Aware Reinforcement Learning for Aggressive Quadrotor Flight

Agile flight tasks such as drone racing and pursuit-evasion require strong acceleration and precise turns, but the available thrust changes as the battery discharges and voltage drops under load. Conservative command limits make this variation easier to tolerate, at the cost of unused performance. We investigate how le...

Alejandro Sánchez Roncero, Olov Andersson, Petter Ogren · 0 citations
Preprint Aug 2026

BVR Sim: An Open and High-Throughput Environment for Heterogeneous Air-Combat Reinforcement Learning

Beyond-visual-range (BVR) air combat is a challenging reinforcement-learning domain characterized by partial observability, long-horizon decision making, energy management, and limited weapons. We present BVR Sim, an open-source Gymnasium-style environment designed for heterogeneous air-combat reinforcement learning. B...

Hao-Cheng Sun, Mu-Lai Tan · 0 citations
Preprint Sep 2026

PINGU: Extending Air-Bearing Spacecraft Emulators with Open-Source Actuators and Learned Control for Contact-Rich Proximity Operations

Low-cost planar air-bearing testbeds have matured into a standard proxy for free-flying spacecraft GNC, but they remain largely thruster-only and are rarely equipped for contact-rich, inertia-coupled manipulation. Building on the open-source ATMOS testbed, we contribute a reaction wheel and two force/torque-sensed robo...

Ricard M. Castan, Akiyoshi Uchida, Aman Arora et al. · 0 citations
Open access Sep 2026

AutoRL: A Tightly Synchronized ROS2–Gazebo Pipeline for Offline-Trained Reinforcement Learning-Based Multirotor Attitude Control

This paper proposes AutoRL, a high-fidelity Robot Operating System 2 (ROS2)–Gazebo simulation pipeline that addresses a critical reproducibility gap in learning-based flight control: existing reinforcement learning (RL) frameworks for unmanned aerial vehicles (UAVs) lack deterministic, step-level coupling between contr...

Khaled Jarrah, O. Rawashdeh · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.