Reprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs Sampling - 专知论文

会员服务 ·

0

吉布斯采样/吉布斯抽样 · CoT · Automator · 样本 · Performer ·

2023 年 5 月 17 日

Reprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs Sampling

翻译：Reprompting：通过吉布斯采样的自动思维链提示推断

Weijia Xu,Andrzej Banburski-Fahey,Nebojsa Jojic

We introduce Reprompting, an iterative sampling algorithm that searches for the Chain-of-Thought (CoT) recipes for a given task without human intervention. Through Gibbs sampling, we infer CoT recipes that work consistently well for a set of training samples. Our method iteratively samples new recipes using previously sampled solutions as parent prompts to solve other training problems. On five Big-Bench Hard tasks that require multi-step reasoning, Reprompting achieves consistently better performance than the zero-shot, few-shot, and human-written CoT baselines. Reprompting can also facilitate transfer of knowledge from a stronger model to a weaker model leading to substantially improved performance of the weaker model. Overall, Reprompting brings up to +17 point improvements over the previous state-of-the-art method that uses human-written CoT prompts.

翻译：我们提出Reprompting，一种无需人工干预即可为给定任务搜索思维链（Chain-of-Thought, CoT）配方的迭代采样算法。通过吉布斯采样，我们推断出一组训练样本上表现稳定良好的CoT配方。该方法迭代地使用先前采样的解作为父提示来采样新配方，以解决其他训练问题。在五项需要多步推理的Big-Bench Hard任务上，Reprompting始终优于零样本、少样本及人工编写的CoT基线方法。Reprompting还可促进知识从更强模型向较弱模型的迁移，从而显著提升较弱模型的性能。总体而言，Reprompting相较于先前使用人工编写CoT提示的最先进方法带来了高达17个百分点的改进。

0

相关内容

吉布斯采样/吉布斯抽样

吉布斯采样/吉布斯抽样

百篇论文纵览大型语言模型最新研究进展

百篇论文纵览大型语言模型最新研究进展

专知会员服务

70+阅读 · 2023年3月31日

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

专知会员服务

138+阅读 · 2022年2月6日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

167+阅读 · 2020年3月18日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

164+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

80+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

106+阅读 · 2019年10月9日

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

图与推荐

2+阅读 · 2022年11月2日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

44+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

【推荐】GAN架构入门综述(资源汇总)

【推荐】GAN架构入门综述(资源汇总)

机器学习研究会

10+阅读 · 2017年9月3日

精确靶向乳腺癌患者的个体化药物研究

国家自然科学基金

0+阅读 · 2015年12月31日

细胞外基质修饰的组织工程神经移植物修复周围神经缺损的研究

国家自然科学基金

0+阅读 · 2014年12月31日

靶向调控SDF-1/CXCR4信号通路干预关节软骨退变的分子机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

动脉粥样硬化易损斑块光声分子显像与治疗基础研究

国家自然科学基金

0+阅读 · 2014年12月31日

高速高精度少自由度并联机器人动力学鲁棒控制研究

国家自然科学基金

0+阅读 · 2013年12月31日

新型多功能GO/HA基仿生材料的构建与性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

新型微孔金属-有机膦酸材料的合成及催化性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

硫醇-烯烃点击可控制备多功能N-P骨架超支化环氧树脂及性能

国家自然科学基金

0+阅读 · 2012年12月31日

功能化纳米结构材料在乏燃料后处理中的应用基础研究

国家自然科学基金

0+阅读 · 2012年12月31日

上消化道癌症原位早期诊断激光拉曼光谱系统的研制

国家自然科学基金

0+阅读 · 2009年12月31日

Improved sampling via learned diffusions

Arxiv

0+阅读 · 2023年7月3日

Exploring Diffusion Models for Unsupervised Video Anomaly Detection

Arxiv

0+阅读 · 2023年7月2日

Robotic Skill Acquisition via Instruction Augmentation with Vision-Language Models

Arxiv

0+阅读 · 2023年7月1日

Convex Optimization in Legged Robots

Arxiv

0+阅读 · 2023年6月30日

Class-Incremental Learning using Diffusion Model for Distillation and Replay

Arxiv

0+阅读 · 2023年6月30日

Learning Agile Flights through Narrow Gaps with Varying Angles using Onboard Sensing

Arxiv

0+阅读 · 2023年6月30日

Diffusion Models in Vision: A Survey

Arxiv

30+阅读 · 2022年9月10日

Prompt Distribution Learning

Arxiv

14+阅读 · 2022年5月6日

Multimodal Categorization of Crisis Events in Social Media

Multimodal Categorization of Crisis Events in Social Media

Arxiv

20+阅读 · 2020年4月10日

Graph Convolutional Label Noise Cleaner: Train a Plug-and-play Action Classifier for Anomaly Detection

Graph Convolutional Label Noise Cleaner: Train a Plug-and-play Action Classifier for Anomaly Detection

Arxiv

15+阅读 · 2019年3月18日

VIP会员

文章信息

相关主题

吉布斯采样/吉布斯抽样

最新内容

《无人系统互操作性导论——无人系统联合架构（JAUS）》

《无人系统互操作性导论——无人系统联合架构（JAUS）》

专知会员服务

7+阅读 · 今天5:53

美空军新型反无人机部队初探

美空军新型反无人机部队初探

专知会员服务

3+阅读 · 今天5:45

《对抗性电磁环境下远程巡飞弹作战的安全指挥与控制数据链》

《对抗性电磁环境下远程巡飞弹作战的安全指挥与控制数据链》

专知会员服务

2+阅读 · 今天5:23

《北约下一代建模与仿真（NexGen M&S）计划》2026年69页

《北约下一代建模与仿真（NexGen M&S）计划》2026年69页

专知会员服务

1+阅读 · 今天5:11

《防空交战流程的概率建模研究》

《防空交战流程的概率建模研究》

专知会员服务

6+阅读 · 今天5:04

ICML 2026 教程 | 数值优化理论还重要吗？

ICML 2026 教程 | 数值优化理论还重要吗？

专知会员服务

4+阅读 · 7月26日

ICM 2026 | 陶哲轩：人工智能时代的数学

ICM 2026 | 陶哲轩：人工智能时代的数学

专知会员服务

7+阅读 · 7月26日

《面向可扩展高韧性无人机集群网络的速度感知分层通信框架》

《面向可扩展高韧性无人机集群网络的速度感知分层通信框架》

专知会员服务

8+阅读 · 7月26日

《面向概率推理的可定制战术引擎及其在军事任务规划中的应用》

《面向概率推理的可定制战术引擎及其在军事任务规划中的应用》

专知会员服务

9+阅读 · 7月26日

《先进防空系统选型战略框架：基于巴基斯坦的实证启示》

《先进防空系统选型战略框架：基于巴基斯坦的实证启示》

专知会员服务

8+阅读 · 7月26日

《反无人机交战场景下的战斗归零研究》

《反无人机交战场景下的战斗归零研究》

专知会员服务

7+阅读 · 7月26日

霍尔木兹与不对称作战时代：水雷、无人系统与海军力量的重新定义

霍尔木兹与不对称作战时代：水雷、无人系统与海军力量的重新定义

专知会员服务

4+阅读 · 7月26日

博士论文 | 用代码结构感知方法推进代码大模型

博士论文 | 用代码结构感知方法推进代码大模型

专知会员服务

5+阅读 · 7月25日

综述 | 遥感多模态大模型：领域专用还是通用模型？

综述 | 遥感多模态大模型：领域专用还是通用模型？

专知会员服务

5+阅读 · 7月25日

《面向指挥控制训练与实时北约兼容数据分发的战术模拟器》

《面向指挥控制训练与实时北约兼容数据分发的战术模拟器》

专知会员服务

5+阅读 · 7月25日

相关VIP内容

百篇论文纵览大型语言模型最新研究进展

百篇论文纵览大型语言模型最新研究进展

专知会员服务

70+阅读 · 2023年3月31日

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

【干货书】机器学习设计模式，408页pdf，Machine Learning Design Patterns

专知会员服务

138+阅读 · 2022年2月6日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

167+阅读 · 2020年3月18日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

164+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

80+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

106+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

美空军新型反无人机部队初探

《北约下一代建模与仿真（NexGen M&S）计划》2026年69页

《无人系统互操作性导论——无人系统联合架构（JAUS）》

《对抗性电磁环境下远程巡飞弹作战的安全指挥与控制数据链》

相关资讯

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

图与推荐

2+阅读 · 2022年11月2日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

44+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

【推荐】GAN架构入门综述(资源汇总)

【推荐】GAN架构入门综述(资源汇总)

机器学习研究会

10+阅读 · 2017年9月3日

相关论文

Improved sampling via learned diffusions

Arxiv

0+阅读 · 2023年7月3日

Exploring Diffusion Models for Unsupervised Video Anomaly Detection

Arxiv

0+阅读 · 2023年7月2日

Robotic Skill Acquisition via Instruction Augmentation with Vision-Language Models

Arxiv

0+阅读 · 2023年7月1日

Convex Optimization in Legged Robots

Arxiv

0+阅读 · 2023年6月30日

Class-Incremental Learning using Diffusion Model for Distillation and Replay

Arxiv

0+阅读 · 2023年6月30日

Learning Agile Flights through Narrow Gaps with Varying Angles using Onboard Sensing

Arxiv

0+阅读 · 2023年6月30日

Diffusion Models in Vision: A Survey

Arxiv

30+阅读 · 2022年9月10日

Prompt Distribution Learning

Arxiv

14+阅读 · 2022年5月6日

Multimodal Categorization of Crisis Events in Social Media

Multimodal Categorization of Crisis Events in Social Media

Arxiv

20+阅读 · 2020年4月10日

Graph Convolutional Label Noise Cleaner: Train a Plug-and-play Action Classifier for Anomaly Detection

Graph Convolutional Label Noise Cleaner: Train a Plug-and-play Action Classifier for Anomaly Detection

Arxiv

15+阅读 · 2019年3月18日

相关基金

精确靶向乳腺癌患者的个体化药物研究

国家自然科学基金

0+阅读 · 2015年12月31日

细胞外基质修饰的组织工程神经移植物修复周围神经缺损的研究

国家自然科学基金

0+阅读 · 2014年12月31日

靶向调控SDF-1/CXCR4信号通路干预关节软骨退变的分子机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

动脉粥样硬化易损斑块光声分子显像与治疗基础研究

国家自然科学基金

0+阅读 · 2014年12月31日

高速高精度少自由度并联机器人动力学鲁棒控制研究

国家自然科学基金

0+阅读 · 2013年12月31日

新型多功能GO/HA基仿生材料的构建与性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

新型微孔金属-有机膦酸材料的合成及催化性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

硫醇-烯烃点击可控制备多功能N-P骨架超支化环氧树脂及性能

国家自然科学基金

0+阅读 · 2012年12月31日

功能化纳米结构材料在乏燃料后处理中的应用基础研究

国家自然科学基金

0+阅读 · 2012年12月31日

上消化道癌症原位早期诊断激光拉曼光谱系统的研制

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员