Research on UAV Anti-UAV Strategies Based on Deep Reinforcement Learning
This paper constructs a reinforcement learning framework based on the PPO algorithm for drone air combat to solve 1v1 pursuit-evasion in 2D beyond-visual-range air combat. Firstly, the mission scenario is modeled, defining key roles of ATA and AA. Then, state transition models of pursuer and evader are built based on f...