<p>Scalability is a key factor for the deployment of multi-agent reinforcement learning (MARL), and how to effectively control the computational cost while improving scalability has become a core challenge in this field. Although existing research has extensively explored how to enhance the scalability of algorithms, it generally overlooks the computational overhead and training costs of models. Addressing this critical challenge, we introduce the Efficient Evolutionary Curriculum Learning (E2CL). This method integrates evolutionary ideas into curriculum learning, training multiple populations of agents in parallel at each stage. Through competitive, selection, crossover, and mutation operations, the fittest populations are selected for training in the next stage, thereby overcoming the bottleneck issue in traditional curriculum learning where the optimal agent from the previous stage struggles to adapt to new stages. To further optimize training efficiency, we propose a dual-dimensional hybrid probability sampling mechanism for population selection based on population rewards and training stability. This mechanism effectively reduces redundant competition during the evolutionary process, significantly lowering the computational overhead of the algorithm. Additionally, we also design a hybrid curriculum knowledge transfer method that combines model reload and buffer reuse to maximize the utilization of knowledge between curricula, enhancing the Jumpstart performance in new curriculum stages. Simulation results in cooperative and adversarial tasks show that E2CL significantly reduces the training cost while ensuring performance, thereby establishing a balance between performance and resource consumption. E2CL offers an efficient paradigm for multi-agent collaborative training and is a practical solution for MARL training in resource-sensitive scenarios.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Efficient evolutionary curriculum learning for scalable multi-agent reinforcement learning

  • Chao Li,
  • Yanfei Liu,
  • Jieling Wang,
  • Zhong Wang,
  • Chengjin Wang,
  • Qi Tian

摘要

Scalability is a key factor for the deployment of multi-agent reinforcement learning (MARL), and how to effectively control the computational cost while improving scalability has become a core challenge in this field. Although existing research has extensively explored how to enhance the scalability of algorithms, it generally overlooks the computational overhead and training costs of models. Addressing this critical challenge, we introduce the Efficient Evolutionary Curriculum Learning (E2CL). This method integrates evolutionary ideas into curriculum learning, training multiple populations of agents in parallel at each stage. Through competitive, selection, crossover, and mutation operations, the fittest populations are selected for training in the next stage, thereby overcoming the bottleneck issue in traditional curriculum learning where the optimal agent from the previous stage struggles to adapt to new stages. To further optimize training efficiency, we propose a dual-dimensional hybrid probability sampling mechanism for population selection based on population rewards and training stability. This mechanism effectively reduces redundant competition during the evolutionary process, significantly lowering the computational overhead of the algorithm. Additionally, we also design a hybrid curriculum knowledge transfer method that combines model reload and buffer reuse to maximize the utilization of knowledge between curricula, enhancing the Jumpstart performance in new curriculum stages. Simulation results in cooperative and adversarial tasks show that E2CL significantly reduces the training cost while ensuring performance, thereby establishing a balance between performance and resource consumption. E2CL offers an efficient paradigm for multi-agent collaborative training and is a practical solution for MARL training in resource-sensitive scenarios.