错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Algorithms for Learning Value-Aligned Policies Considering Admissibility Relaxation

  • Andrés Holgado-Sánchez,
  • Joaquín Arias,
  • Holger Billhardt,
  • Sascha Ossowski

摘要

The emerging field of value awareness engineering claims that software agents and systems should be value-aware, i.e. they must make decisions in accordance with human values. In this context, such agents must be capable of explicitly reasoning as to how far different courses of action are aligned with these values. For this purpose, values are often modelled as preferences over states or actions, which are then aggregated to determine the sequences of actions that are maximally aligned with a certain value. Recently, additional value admissibility constraints at this level have been considered as well.