The increase in the number of conversational systems (chatbots) applied to different scenarios in society is notable. However, the development of metrics and evaluation methods for chatbots still remains an open line of research. The aim is to create evaluation methods that are less and less invasive and that do not reduce the participation of users or other human agents. In this work, in the methods section, a systematic review is carried out on different evaluation methods for chatbots. Then, in the same section, new metrics for the evaluation of conversations with chatbots inspired by the neutrosophic theory are presented. In the results section, the validation of the proposed model is carried out. The applicability of the model is evaluated and the proposal is subjected to expert triangulation methods. In the work, it is possible to demonstrate that the application of neutrosophic logic can contribute to achieving a more natural response from chatbots. This effect can be useful to mitigate the presence of false responses or hallucinations.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Measurement of Perceived Quality in Conversational Systems (Chatbots)

  • Luis Gabriel Hernández Pupo,
  • Raykenler Yzquierdo Herrera,
  • Luis Alvarado Acuña,
  • Pedro Yobanis Piñero Pérez,
  • Iliana Pérez Pupo,
  • Rafael Bello Pérez

摘要

The increase in the number of conversational systems (chatbots) applied to different scenarios in society is notable. However, the development of metrics and evaluation methods for chatbots still remains an open line of research. The aim is to create evaluation methods that are less and less invasive and that do not reduce the participation of users or other human agents. In this work, in the methods section, a systematic review is carried out on different evaluation methods for chatbots. Then, in the same section, new metrics for the evaluation of conversations with chatbots inspired by the neutrosophic theory are presented. In the results section, the validation of the proposed model is carried out. The applicability of the model is evaluated and the proposal is subjected to expert triangulation methods. In the work, it is possible to demonstrate that the application of neutrosophic logic can contribute to achieving a more natural response from chatbots. This effect can be useful to mitigate the presence of false responses or hallucinations.