With the rise of generative AI models, such as large language models (LLMs), in educational settings, there is a growing demand to ensure the quality of AI-generated multiple-choice questions (MCQs) used in higher education. Traditional quiz development methods fall short in addressing the unique challenges posed by AI-generated content, such as consistency, cognitive demand, and question uniqueness. This paper presents the QUEST framework, a structured approach designed specifically to evaluate the quality of LLM-generated MCQs across five dimensions: Quality, Uniqueness, Effort, Structure, and Transparency. Following an iterative research process, AI-generated questions were assessed and refined using QUEST, revealing that the framework effectively improves question clarity, relevance, and educational value. The findings suggest that QUEST is a viable tool for educators to maintain high-quality standards in AI-generated assessments, ensuring these resources meet the pedagogical needs of diverse learners in higher education.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Ensuring Quality in AI-Generated Multiple-Choice Questions for Higher Education with the QUEST Framework

  • Martin Ebner,
  • Benedikt Brünner,
  • Noel Forjan,
  • Sandra Schön

摘要

With the rise of generative AI models, such as large language models (LLMs), in educational settings, there is a growing demand to ensure the quality of AI-generated multiple-choice questions (MCQs) used in higher education. Traditional quiz development methods fall short in addressing the unique challenges posed by AI-generated content, such as consistency, cognitive demand, and question uniqueness. This paper presents the QUEST framework, a structured approach designed specifically to evaluate the quality of LLM-generated MCQs across five dimensions: Quality, Uniqueness, Effort, Structure, and Transparency. Following an iterative research process, AI-generated questions were assessed and refined using QUEST, revealing that the framework effectively improves question clarity, relevance, and educational value. The findings suggest that QUEST is a viable tool for educators to maintain high-quality standards in AI-generated assessments, ensuring these resources meet the pedagogical needs of diverse learners in higher education.