ChatGPT is built on a large language model trained on an enormous corpus of human text to emulate human conversation. Despite lacking any explicit programming regarding the laws of physics, recent work has demonstrated that GPT-3.5 could pass an introductory physics course at some nominal level and register something close to a minimal understanding of Newtonian Mechanics on the Force Concept Inventory. This work replicates those results and also demonstrates that the latest version, GPT-4, has reached a much higher mark in the latter context. Indeed, its responses come quite close to perfectly demonstrating expert-level competence, with a few very notable exceptions and limitations. We briefly comment on the implications of this for the future of physics education and pedagogy.
翻译:ChatGPT基于一个大型语言模型构建,该模型在庞大的文本语料库上训练以模拟人类对话。尽管缺乏任何关于物理定律的显式编程,近期研究表明GPT-3.5能在名义水平上通过物理入门课程,并在"力概念量表"上表现出接近牛顿力学的最低理解。本研究复现了这些结果,并证明最新版本GPT-4在后一测试中取得了显著更高的分数。事实上,其回答几乎完美展现出专家级能力,仅有少数显著例外和局限。我们简要评述了这对物理教育及教学法未来的启示。