伦理引擎：用于大型语言模型可及性心理测量评估的模块化流程 (The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models)

As Large Language Models increasingly mediate human communication and decision-making, understanding their value expression becomes critical for research across disciplines. This work presents the Ethics Engine, a modular Python pipeline that transforms psychometric assessment of LLMs from a technically complex endeavor into an accessible research tool. The pipeline demonstrates how thoughtful infrastructure design can expand participation in AI research, enabling investigators across cognitive science, political psychology, education, and other fields to study value expression in language models. Recent adoption by University of Edinburgh researchers studying authoritarianism validates its research utility, processing over 10,000 AI responses across multiple models and contexts. We argue that such tools fundamentally change the landscape of AI research by lowering technical barriers while maintaining scientific rigor. As LLMs increasingly serve as cognitive infrastructure, their embedded values shape millions of daily interactions. Without systematic measurement of these value expressions, we deploy systems whose moral influence remains uncharted. The Ethics Engine enables the rigorous assessment necessary for informed governance of these influential technologies.

翻译：随着大型语言模型日益介入人类沟通与决策过程，理解其价值表达已成为跨学科研究的关键课题。本研究提出伦理引擎——一个模块化的Python流程，将LLMs的心理测量评估从技术复杂的任务转化为可及的研究工具。该流程展示了深思熟虑的基础设施设计如何拓展人工智能研究的参与度，使认知科学、政治心理学、教育学等领域的学者能够研究语言模型中的价值表达。爱丁堡大学研究者在威权主义研究中对该工具的近期应用验证了其研究效用，已处理超过10,000份跨多模型与多情境的AI响应。我们认为此类工具通过降低技术门槛同时保持科学严谨性，正在从根本上改变人工智能研究的格局。随着LLMs日益成为认知基础设施，其内嵌价值观正塑造着数百万次的日常交互。若缺乏对这些价值表达的系统性测量，我们将部署道德影响未知的系统。伦理引擎为实现这些影响深远技术的知情治理提供了必要的严谨评估手段。

相关内容

MoDELS

关注 45

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

O’Reilly报告：知识图谱崛起——面向现代数据集成和数据结构体系，“The Rise of the Knowledge Graph——Toward Modern Data Integration and the Data Fabric Architecture”

专知会员服务

49+阅读 · 2022年2月18日

【CHI2020-微软】解释可解释性:理解数据科学家使用机器学习的可解释性工具，Interpreting Interpretability: Understanding Data Scientists’Use of Interpretability Tools for Machine Learning

专知会员服务

55+阅读 · 2020年3月8日

FlowQA: Grasping Flow in History for Conversational Machine Comprehension

专知会员服务

34+阅读 · 2019年10月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日