This paper presents a performance comparison of three large language models (LLMs), namely OpenAI ChatGPT, Microsoft Bing Chat (BingChat), and Google Bard, on the VNHSGE English dataset. The performance of BingChat, Bard, and ChatGPT (GPT-3.5) is 92.4\%, 86\%, and 79.2\%, respectively. The results show that BingChat is better than ChatGPT and Bard. Therefore, BingChat and Bard can replace ChatGPT while ChatGPT is not yet officially available in Vietnam. The results also indicate that BingChat, Bard and ChatGPT outperform Vietnamese students in English language proficiency. The findings of this study contribute to the understanding of the potential of LLMs in English language education. The remarkable performance of ChatGPT, BingChat, and Bard demonstrates their potential as effective tools for teaching and learning English at the high school level.
翻译:本文对三种大型语言模型(LLMs)——OpenAI ChatGPT、Microsoft Bing Chat(BingChat)和Google Bard——在VNHSGE英语数据集上的性能进行了比较。BingChat、Bard和ChatGPT(GPT-3.5)的性能分别为92.4%、86%和79.2%。结果表明,BingChat优于ChatGPT和Bard。因此,在ChatGPT尚未在越南正式开放的情况下,BingChat和Bard可以替代ChatGPT。结果还显示,BingChat、Bard和ChatGPT在英语语言能力上均优于越南学生。本研究的结果有助于理解LLMs在英语语言教育中的潜力。ChatGPT、BingChat和Bard的卓越表现表明,它们有潜力成为高中英语教学的有效工具。