错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Orbital Multi-Player Pursuit-Evasion Game with Deep Reinforcement Learning

  • Zhen-yu Li,
  • Si Chen,
  • Chenghong Zhou,
  • Wei Sun

摘要

This study investigates an orbital multi-player “encirclement-capture” game where multiple pursuers with an encirclement configuration aim to capture an evader, whereas the evader tries to escape all of the pursuers. First, an elliptic encirclement configuration is designed for the pursuers to exploit the initial position advantage. Then, the pursuit-evasion process for capturing the evader is formulated as a discrete Markov game. To acquire superior pursuit-evasion strategies, a distributed distributional deep deterministic policy gradient algorithm is employed and modified for the multi-player game. The main structure of the algorithm is modified as a parallel adversarial-learning framework to achieve efficient two-sided training. Meanwhile, the policy networks and policy-gradient calculation are modified to achieve a decentralized-decision coordination among multiple pursuers. Simulations showed that the pursuers and evader trained via the proposed algorithm can learn ‘active-cooperation’ pursuit strategy and ‘multi-target’ evasion strategy, respectively. Meanwhile, the obtained strategies outperform traditional pursuit-evasion strategies.