<p>Artificial intelligence (AI) chatbots show remarkable abilities across applications. Despite a growing literature, their capability in the field of entrepreneurship is not fully understood. The aim of this study is to empirically evaluate and compare capabilities of five major AI chatbots—GPT-3.5, GPT-4, Gemini 1.0, Llama 2, and Claude—in the context of entrepreneurship theory, using a benchmark entrepreneurship test. In particular, the performance of the chatbots on a set of multiple-choice questions, short-answer questions, and essay questions related to entrepreneurship is assessed. The results indicate that GPT-4 delivers the strongest overall performance. Meanwhile, Llama 2 offers precise responses with a significantly lower word count compared to the GPT models. Although chatbots do not always provide correct or precise answers to questions or complex prompts, they still prove to be valuable analytical tools for entrepreneurs. While the study offers compelling insights into chatbots’ grasp of entrepreneurship concepts, the findings are somewhat limited by the scarce availability of data.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Comparative analysis of leading artificial intelligence chatbots in the context of entrepreneurship

  • Firuz Kamalov,
  • David Santandreu Calonge,
  • Patrik T. Hultberg,
  • Linda Smail,
  • Dima Jamali

摘要

Artificial intelligence (AI) chatbots show remarkable abilities across applications. Despite a growing literature, their capability in the field of entrepreneurship is not fully understood. The aim of this study is to empirically evaluate and compare capabilities of five major AI chatbots—GPT-3.5, GPT-4, Gemini 1.0, Llama 2, and Claude—in the context of entrepreneurship theory, using a benchmark entrepreneurship test. In particular, the performance of the chatbots on a set of multiple-choice questions, short-answer questions, and essay questions related to entrepreneurship is assessed. The results indicate that GPT-4 delivers the strongest overall performance. Meanwhile, Llama 2 offers precise responses with a significantly lower word count compared to the GPT models. Although chatbots do not always provide correct or precise answers to questions or complex prompts, they still prove to be valuable analytical tools for entrepreneurs. While the study offers compelling insights into chatbots’ grasp of entrepreneurship concepts, the findings are somewhat limited by the scarce availability of data.