错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Personalization for web-based services using offline reinforcement learning

  • Pavlos Athanasios Apostolopoulos,
  • Zehui Wang,
  • Hanson Wang,
  • Tenghyu Xu,
  • Chad Zhou,
  • Kittipate Virochsiri,
  • Norm Zhou,
  • Igor L. Markov

摘要

Large-scale Web-based services present opportunities for improving UI policies based on observed user interactions. We address challenges of learning such policies through offline reinforcement learning (RL). Deployed in a production system for user authentication in a major social network, it significantly improves long-term objectives. We articulate practical challenges, provide insights on training and evaluation of offline RL, and discuss generalizations toward offline RL’s deployment in industry-scale applications.