Beyond Text Following: Repairable Arbitration Reversals in Audio-Language Models

Audio-language models (ALMs) often follow text that conflicts with audio, even when the audio evidence is clear. This raises a basic question: is the audio-supported answer unavailable, or is it represented but overridden by the conflicting text? We examine this question using a same-audio counterfactual that keeps the audio fixed, removes only the conflicting text, and measures the resulting shift in model preference. Across five ALMs and four conflict tasks, 64.1% of conflict samples show a sign flip: the same-audio branch prefers the audio-supported answer, whereas the joint branch prefers the text-supported answer. This pattern suggests that the relevant audio evidence is encoded but loses in arbitration. Activation patching further localizes the reversal to answer-position computation, and patching effects closely track output candidate-score differences (Spearman rho=0.93). Using this diagnostic, we propose Gated Audio Counterfactual Logit Correction (GACL), a training-free decoding rule that interpolates between joint and same-audio scores. Under a strict 5 pp faithfulness-drop budget, GACL improves nAUC by 17.8 points over the best contrastive baseline and transfers without retuning to vision-text arbitration (up to +40.5 pp).

翻译：音频-语言模型（ALMs）常常遵从与音频冲突的文本指令，即便音频证据十分明确。这引出一个基本问题：音频支持的答案是不可获得的，还是已被表征但被冲突文本所覆盖？我们通过使用同音频反事实来研究此问题，该反事实固定音频不变，仅移除冲突文本，并测量模型偏好的相应变化。在五种ALM和四个冲突任务中，64.1%的冲突样本显示出符号翻转：同音频分支倾向于音频支持的答案，而联合分支倾向于文本支持的答案。这一模式表明相关音频证据已被编码，但在仲裁中落败。激活补丁进一步将逆转定位到答案位置的计算，且补丁效应与输出候选分数差异高度相关（Spearman rho=0.93）。基于这一诊断，我们提出门控音频反事实对数几率修正（GACL），一种无需训练的解码规则，用于插值联合分数与同音频分数。在严格5个百分点（pp）的忠实度下降预算下，GACL在nAUC指标上比最优对比基线提升17.8点，并可零调参迁移至视觉-文本仲裁（提升高达40.5 pp）。

相关内容

MoDELS

关注 45

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

【综述】大型音频语言模型综述：泛化、可信与未来展望

专知会员服务

12+阅读 · 5月21日

【CIKM2025教程】语言模型的公平性：一篇教程，170页ppt

专知会员服务

16+阅读 · 2025年11月16日

《语音大语言模型》最新进展综述

专知会员服务

58+阅读 · 2024年10月8日

【博士论文】语言模型与人类偏好对齐，148页pdf

专知会员服务

32+阅读 · 2024年4月21日