错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Analyzing the Efficacy of Large Language Models: A Comparative Study

  • Sonia Khetarpaul,
  • Dolly Sharma,
  • Shreya Sinha,
  • Aryan Nagpal,
  • Aarush Narang

摘要

With the rise of large language models (LLMs) and natural language processing (NLP) methods in businesses and industries, our research evaluates two prominent LLMs: GPT-3.5 Turbo by OpenAI and Llama 2 by Meta. We used an automated process to generate and fine-tune question-answer pairs, enhancing accuracy. Using established metrics, we quantified performance and developed a comprehensive evaluation metric. Our analysis highlighted the need for further improvements to address prevalent issues such as inaccuracies and ambiguities.