不要为工具适配小型语言模型；让工具模式适配模型 (Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models) - 专知论文

会员服务 ·

0

工具 · 预训练 · 适配 · 对齐 · 小型语言模型 ·

Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models

翻译：不要为工具适配小型语言模型；让工具模式适配模型

Jonggeun Lee,Woojung Song,Jongwook Han,Haesung Pyun,Yohan Jo

from arxiv, 22 pages

Small language models (SLMs) enable scalable multi-agent tool systems where multiple SLMs handle subtasks orchestrated by a powerful coordinator. However, they struggle with tool-use tasks, particularly in selecting appropriate tools and identifying correct parameters. A common failure mode is schema misalignment: models hallucinate plausible but nonexistent tool names that reflect naming conventions internalized during pretraining but absent from the provided tool schema. Rather than forcing models to adapt to arbitrary schemas, we propose adapting schemas to align with models' pretrained knowledge. We introduce PA-Tool (Pretraining-Aligned Tool Schema Generation), a training-free method that leverages peakedness, a signal from contamination detection indicating pretraining familiarity, to rename tool components. By generating multiple candidates and selecting those with the highest peakedness across samples, PA-Tool identifies pretraining-aligned naming patterns. Experiments on MetaTool and RoTBench show improvements of up to 17%, with schema misalignment errors reduced by 80%. PA-Tool enables small models to approach state-of-the-art performance while maintaining computational efficiency in adapting to new tools without retraining. Our work demonstrates that schema-level interventions can unlock the tool-use potential of resource-efficient models by adapting schemas to models rather than models to schemas.

翻译：小型语言模型（SLMs）能够实现可扩展的多智能体工具系统，其中多个SLM处理由强大协调器编排的子任务。然而，它们在工具使用任务上存在困难，特别是在选择合适的工具和识别正确参数方面。一种常见的失败模式是模式失配：模型会幻觉出看似合理但实际不存在的工具名称，这些名称反映了预训练期间内化的命名惯例，但在提供的工具模式中并不存在。我们提出，与其强迫模型适应任意模式，不如调整模式以对齐模型的预训练知识。我们引入PA-Tool（预训练对齐工具模式生成），这是一种无需训练的方法，它利用峰值度——一种来自污染检测的信号，指示预训练的熟悉程度——来重命名工具组件。通过生成多个候选名称并选择在样本中具有最高峰值度的名称，PA-Tool识别出与预训练对齐的命名模式。在MetaTool和RoTBench上的实验显示性能提升高达17%，模式失配错误减少了80%。PA-Tool使小型模型能够接近最先进的性能，同时在适应新工具时保持计算效率，无需重新训练。我们的工作表明，通过让模式适应模型而非模型适应模式，模式层面的干预能够释放资源高效模型的工具使用潜力。

0

相关内容

多模态大语言模型下游调优中“保持自我”的重要性

多模态大语言模型下游调优中“保持自我”的重要性

专知会员服务

17+阅读 · 2025年12月15日

面向性能、成本效益、云边隐私与可信性的大小语言模型协作综述

面向性能、成本效益、云边隐私与可信性的大小语言模型协作综述

专知会员服务

15+阅读 · 2025年10月18日

在无标注条件下适配视觉—语言模型：全面综述

在无标注条件下适配视觉—语言模型：全面综述

专知会员服务

13+阅读 · 2025年8月9日

大语言模型与小语言模型协同机制综述

大语言模型与小语言模型协同机制综述

专知会员服务

38+阅读 · 2025年5月15日

中文版（万字长文） | 小型语言模型：紧凑型AI的崛起与微软Phi系列

中文版（万字长文） | 小型语言模型：紧凑型AI的崛起与微软Phi系列

专知会员服务

13+阅读 · 2025年5月9日

【新书】设计大型语言模型应用：一种面向LLMs的整体方法

【新书】设计大型语言模型应用：一种面向LLMs的整体方法

专知会员服务

55+阅读 · 2025年3月16日

小型语言模型综述

小型语言模型综述

专知会员服务

54+阅读 · 2024年10月29日

【CVPR2024】探索多模态大型语言模型中视觉提示的可转移性

【CVPR2024】探索多模态大型语言模型中视觉提示的可转移性

专知会员服务

21+阅读 · 2024年4月18日

大模型的“幻觉”如何克服？腾讯AILab等《大型语言模型中的幻觉》，全面阐述检测、解释和减轻幻觉

大模型的“幻觉”如何克服？腾讯AILab等《大型语言模型中的幻觉》，全面阐述检测、解释和减轻幻觉

专知会员服务

72+阅读 · 2023年9月7日

中科大腾讯最新《多模态大型语言模型》综述，详述多模态指令微调、上下文学习、思维链和辅助视觉推理技术

中科大腾讯最新《多模态大型语言模型》综述，详述多模态指令微调、上下文学习、思维链和辅助视觉推理技术

专知会员服务

105+阅读 · 2023年6月27日

从T5到GPT-4最新最全梳理，人大等《大型语言模型综述》，51页pdf详述大模型进展

从T5到GPT-4最新最全梳理，人大等《大型语言模型综述》，51页pdf详述大模型进展

专知

25+阅读 · 2023年4月4日

绝对干货！NLP预训练模型：从transformer到albert

绝对干货！NLP预训练模型：从transformer到albert

新智元

13+阅读 · 2019年11月10日

模型不work怎么办？141页PPT告诉你怎么改模型

模型不work怎么办？141页PPT告诉你怎么改模型

专知

17+阅读 · 2019年10月31日

预训练语言模型关系图+必读论文列表，清华荣誉出品

预训练语言模型关系图+必读论文列表，清华荣誉出品

机器之心

18+阅读 · 2019年10月11日

最新必读【预训练语言模型(BERT/XLNet等)】论文，Google/微软/华为ICLR2020提交论文

最新必读【预训练语言模型(BERT/XLNet等)】论文，Google/微软/华为ICLR2020提交论文

专知

36+阅读 · 2019年9月29日

你的TextGAN调出来了么？来看看人在怎么调的

你的TextGAN调出来了么？来看看人在怎么调的

专知

85+阅读 · 2019年6月6日

BAM！利用知识蒸馏和多任务学习构建的通用语言模型

BAM！利用知识蒸馏和多任务学习构建的通用语言模型

机器之心

15+阅读 · 2019年3月18日

用模型不确定性理解模型

用模型不确定性理解模型

论智

11+阅读 · 2018年9月5日

NLP通用模型诞生？一个模型搞定十大自然语言常见任务

NLP通用模型诞生？一个模型搞定十大自然语言常见任务

人工智能头条

10+阅读 · 2018年6月29日

自然语言处理中的Attention Model：是什么及为什么

自然语言处理中的Attention Model：是什么及为什么

新智元

11+阅读 · 2017年7月13日

支持新产品快速设计的复杂产品系统功能模块化方法

国家自然科学基金

1+阅读 · 2015年12月31日

近似计算中基于概率图模型的软错误量化方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

基于非独立同分布学习理论的图模型词义消歧及领域适应方法研究

国家自然科学基金

1+阅读 · 2015年12月31日

即时通信中的隐蔽通信模型及方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

方差正则化的分类模型选择方法研究

国家自然科学基金

1+阅读 · 2015年12月31日

基于犹豫模糊语言信息的定性决策理论与方法

国家自然科学基金

2+阅读 · 2015年12月31日

共现潜在语义向量空间模型及其语义核的构建与应用研究

国家自然科学基金

1+阅读 · 2015年12月31日

普适计算对象感知多模态不精确性数据融合算法研究

国家自然科学基金

5+阅读 · 2014年12月31日

面向地理模型集成与运行的数据适配方法研究

国家自然科学基金

1+阅读 · 2014年12月31日

基于群体智能的多Agent协作模型与适应性研究

国家自然科学基金

18+阅读 · 2009年12月31日

TASTE: Text-Aligned Speech Tokenization and Embedding for Spoken Language Modeling

Arxiv

0+阅读 · 2月5日

KVSmooth: Mitigating Hallucination in Multi-modal Large Language Models through Key-Value Smoothing

Arxiv

0+阅读 · 2月4日

Language Models Struggle to Use Representations Learned In-Context

Arxiv

0+阅读 · 2月4日

Programming Language Confusion: When Code LLMs Can't Keep their Languages Straight

Arxiv

0+阅读 · 2月2日

Knowing the Facts but Choosing the Shortcut: Understanding How Large Language Models Compare Entities

Arxiv

0+阅读 · 1月24日

Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks

Arxiv

0+阅读 · 1月18日

Seeing Right but Saying Wrong: Inter- and Intra-Layer Refinement in MLLMs without Training

Arxiv

0+阅读 · 1月12日

Big Reasoning with Small Models: Instruction Retrieval at Inference Time

Arxiv

0+阅读 · 1月7日

Fine-tuning Small Language Models as Efficient Enterprise Search Relevance Labelers

Fine-tuning Small Language Models as Efficient Enterprise Search Relevance Labelers

Arxiv

0+阅读 · 1月6日

Protecting multimodal large language models against misleading visualizations

Arxiv

0+阅读 · 1月6日

VIP会员

文章信息

相关主题

小型语言模型

相关VIP内容

多模态大语言模型下游调优中“保持自我”的重要性

多模态大语言模型下游调优中“保持自我”的重要性

专知会员服务

17+阅读 · 2025年12月15日

面向性能、成本效益、云边隐私与可信性的大小语言模型协作综述

面向性能、成本效益、云边隐私与可信性的大小语言模型协作综述

专知会员服务

15+阅读 · 2025年10月18日

在无标注条件下适配视觉—语言模型：全面综述

在无标注条件下适配视觉—语言模型：全面综述

专知会员服务

13+阅读 · 2025年8月9日

大语言模型与小语言模型协同机制综述

大语言模型与小语言模型协同机制综述

专知会员服务

38+阅读 · 2025年5月15日

中文版（万字长文） | 小型语言模型：紧凑型AI的崛起与微软Phi系列

中文版（万字长文） | 小型语言模型：紧凑型AI的崛起与微软Phi系列

专知会员服务

13+阅读 · 2025年5月9日

【新书】设计大型语言模型应用：一种面向LLMs的整体方法

【新书】设计大型语言模型应用：一种面向LLMs的整体方法

专知会员服务

55+阅读 · 2025年3月16日

小型语言模型综述

小型语言模型综述

专知会员服务

54+阅读 · 2024年10月29日

【CVPR2024】探索多模态大型语言模型中视觉提示的可转移性

【CVPR2024】探索多模态大型语言模型中视觉提示的可转移性

专知会员服务

21+阅读 · 2024年4月18日

大模型的“幻觉”如何克服？腾讯AILab等《大型语言模型中的幻觉》，全面阐述检测、解释和减轻幻觉

大模型的“幻觉”如何克服？腾讯AILab等《大型语言模型中的幻觉》，全面阐述检测、解释和减轻幻觉

专知会员服务

72+阅读 · 2023年9月7日

中科大腾讯最新《多模态大型语言模型》综述，详述多模态指令微调、上下文学习、思维链和辅助视觉推理技术

中科大腾讯最新《多模态大型语言模型》综述，详述多模态指令微调、上下文学习、思维链和辅助视觉推理技术

专知会员服务

105+阅读 · 2023年6月27日

热门VIP内容

开通专知VIP会员享更多权益服务

《无人机与战争：被忽视的环境影响及无人机保护潜力》

俄罗斯规划未来无人机驱动军队

《整合杀伤链：一个用于边缘目标验证与战术推理的零样本框架》最新资料

《人工智能、武器与影响力：前沿模型在模拟核危机中展现复杂推理》2026最新46页报告

相关资讯

从T5到GPT-4最新最全梳理，人大等《大型语言模型综述》，51页pdf详述大模型进展

从T5到GPT-4最新最全梳理，人大等《大型语言模型综述》，51页pdf详述大模型进展

专知

25+阅读 · 2023年4月4日

绝对干货！NLP预训练模型：从transformer到albert

绝对干货！NLP预训练模型：从transformer到albert

新智元

13+阅读 · 2019年11月10日

模型不work怎么办？141页PPT告诉你怎么改模型

模型不work怎么办？141页PPT告诉你怎么改模型

专知

17+阅读 · 2019年10月31日

预训练语言模型关系图+必读论文列表，清华荣誉出品

预训练语言模型关系图+必读论文列表，清华荣誉出品

机器之心

18+阅读 · 2019年10月11日

最新必读【预训练语言模型(BERT/XLNet等)】论文，Google/微软/华为ICLR2020提交论文

最新必读【预训练语言模型(BERT/XLNet等)】论文，Google/微软/华为ICLR2020提交论文

专知

36+阅读 · 2019年9月29日

你的TextGAN调出来了么？来看看人在怎么调的

你的TextGAN调出来了么？来看看人在怎么调的

专知

85+阅读 · 2019年6月6日

BAM！利用知识蒸馏和多任务学习构建的通用语言模型

BAM！利用知识蒸馏和多任务学习构建的通用语言模型

机器之心

15+阅读 · 2019年3月18日

用模型不确定性理解模型

用模型不确定性理解模型

论智

11+阅读 · 2018年9月5日

NLP通用模型诞生？一个模型搞定十大自然语言常见任务

NLP通用模型诞生？一个模型搞定十大自然语言常见任务

人工智能头条

10+阅读 · 2018年6月29日

自然语言处理中的Attention Model：是什么及为什么

自然语言处理中的Attention Model：是什么及为什么

新智元

11+阅读 · 2017年7月13日

相关论文

TASTE: Text-Aligned Speech Tokenization and Embedding for Spoken Language Modeling

Arxiv

0+阅读 · 2月5日

KVSmooth: Mitigating Hallucination in Multi-modal Large Language Models through Key-Value Smoothing

Arxiv

0+阅读 · 2月4日

Language Models Struggle to Use Representations Learned In-Context

Arxiv

0+阅读 · 2月4日

Programming Language Confusion: When Code LLMs Can't Keep their Languages Straight

Arxiv

0+阅读 · 2月2日

Knowing the Facts but Choosing the Shortcut: Understanding How Large Language Models Compare Entities

Arxiv

0+阅读 · 1月24日

Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks

Arxiv

0+阅读 · 1月18日

Seeing Right but Saying Wrong: Inter- and Intra-Layer Refinement in MLLMs without Training

Arxiv

0+阅读 · 1月12日

Big Reasoning with Small Models: Instruction Retrieval at Inference Time

Arxiv

0+阅读 · 1月7日

Fine-tuning Small Language Models as Efficient Enterprise Search Relevance Labelers

Fine-tuning Small Language Models as Efficient Enterprise Search Relevance Labelers

Arxiv

0+阅读 · 1月6日

Protecting multimodal large language models against misleading visualizations

Arxiv

0+阅读 · 1月6日

相关基金

支持新产品快速设计的复杂产品系统功能模块化方法

国家自然科学基金

1+阅读 · 2015年12月31日

近似计算中基于概率图模型的软错误量化方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

基于非独立同分布学习理论的图模型词义消歧及领域适应方法研究

国家自然科学基金

1+阅读 · 2015年12月31日

即时通信中的隐蔽通信模型及方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

方差正则化的分类模型选择方法研究

国家自然科学基金

1+阅读 · 2015年12月31日

基于犹豫模糊语言信息的定性决策理论与方法

国家自然科学基金

2+阅读 · 2015年12月31日

共现潜在语义向量空间模型及其语义核的构建与应用研究

国家自然科学基金

1+阅读 · 2015年12月31日

普适计算对象感知多模态不精确性数据融合算法研究

国家自然科学基金

5+阅读 · 2014年12月31日

面向地理模型集成与运行的数据适配方法研究

国家自然科学基金

1+阅读 · 2014年12月31日

基于群体智能的多Agent协作模型与适应性研究

国家自然科学基金

18+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员