错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Multi-armed bandit games

  • Kemal Gürsoy

摘要

A sequential optimization model, known as the multi-armed bandit problem, is concerned with optimal allocation of resources between competing activities, in order to generate the most likely benefits, for a given period of time. In this work, following the objective of a multi-armed bandit problem, we consider a mean-field game model to approach to a large number of multi-armed bandit problems, and propose some connections between dynamic games and sequential optimization problems.