This paper explores a methodology that integrates the proportional integral derivative (PID) method with deep reinforcement learning (DRL). Constructing a compensatory controller using DRL to enhance the control performance of a PID controller. The proposed method is utilized in developing a fixed-wing unmanned aerial vehicle (UAV) flight controller to achieve longitudinal flight control. The compensatory controller is constructed using the Deep Deterministic Policy Gradient (DDPG) algorithm, tailored with a state-action space selection specifically designed for UAV dynamics and tracking targets. Meanwhile, the penalty term for tracking error and the sparse reward for completing the goal are introduced, and the construction scheme of the reward function is given. Simulation results demonstrate that the DRL-based compensatory controller can improve control performance when PID controller parameters are not optimally tuned to some extent, effectively eliminating pitch angle tracking overshoot and reducing setting time.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Longitudinal Control and Optimization of Fixed-Wing UAV Based on Deep Reinforcement Learning

  • Haiyang He,
  • Zhengen Zhao,
  • Fei Kong

摘要

This paper explores a methodology that integrates the proportional integral derivative (PID) method with deep reinforcement learning (DRL). Constructing a compensatory controller using DRL to enhance the control performance of a PID controller. The proposed method is utilized in developing a fixed-wing unmanned aerial vehicle (UAV) flight controller to achieve longitudinal flight control. The compensatory controller is constructed using the Deep Deterministic Policy Gradient (DDPG) algorithm, tailored with a state-action space selection specifically designed for UAV dynamics and tracking targets. Meanwhile, the penalty term for tracking error and the sparse reward for completing the goal are introduced, and the construction scheme of the reward function is given. Simulation results demonstrate that the DRL-based compensatory controller can improve control performance when PID controller parameters are not optimally tuned to some extent, effectively eliminating pitch angle tracking overshoot and reducing setting time.