用于扩散模型微调的迭代倾斜方法 (Iterative Tilting for Diffusion Fine-Tuning) - 专知论文

会员服务 ·

0

模型微调 · 微调 · 扩散模型 · 梯度 · 分解 ·

2025 年 12 月 2 日

Iterative Tilting for Diffusion Fine-Tuning

翻译：用于扩散模型微调的迭代倾斜方法

Jean Pachebat,Giovanni Conforti,Alain Durmus,Yazid Janati

from arxiv, 14 pages

We introduce iterative tilting, a gradient-free method for fine-tuning diffusion models toward reward-tilted distributions. The method decomposes a large reward tilt $\exp(λr)$ into $N$ sequential smaller tilts, each admitting a tractable score update via first-order Taylor expansion. This requires only forward evaluations of the reward function and avoids backpropagating through sampling chains. We validate on a two-dimensional Gaussian mixture with linear reward, where the exact tilted distribution is available in closed form.

翻译：我们提出了迭代倾斜方法，这是一种无需梯度的扩散模型微调技术，旨在使模型向奖励倾斜分布对齐。该方法将较大的奖励倾斜项 $\\exp(\\lambda r)$ 分解为 $N$ 个连续的小倾斜步骤，每一步通过一阶泰勒展开获得可处理的分数更新。该方法仅需对奖励函数进行前向计算，避免了在采样链中进行反向传播。我们在具有线性奖励的二维高斯混合模型上进行了验证，该场景下精确的倾斜分布具有闭式解。

0

相关内容

模型微调

【ICML2025】免费的Fisher？通过回收平方梯度累加器近似Fisher信息矩阵

【ICML2025】免费的Fisher？通过回收平方梯度累加器近似Fisher信息矩阵

专知会员服务

12+阅读 · 2025年7月28日

UnHiPPO：面向不确定性的状态空间模型初始化方法

UnHiPPO：面向不确定性的状态空间模型初始化方法

专知会员服务

11+阅读 · 2025年6月6日

【NeurIPS2022】黎曼扩散模型

【NeurIPS2022】黎曼扩散模型

专知会员服务

43+阅读 · 2022年9月15日

NeurIPS 2021 | 寻找用于变分布泛化的隐式因果因子

NeurIPS 2021 | 寻找用于变分布泛化的隐式因果因子

专知会员服务

17+阅读 · 2021年12月7日

【ICML2021】随机傅立叶特征的量化算法

专知会员服务

25+阅读 · 2021年7月31日

【CVPR2021】CausalVAE: 引入因果结构的解耦表示学习

【CVPR2021】CausalVAE: 引入因果结构的解耦表示学习

专知

19+阅读 · 2021年3月28日

【CVPR2021】跨模态检索的概率嵌入

【CVPR2021】跨模态检索的概率嵌入

专知

17+阅读 · 2021年3月2日

图节点嵌入(Node Embeddings)概述，9页pdf

图节点嵌入(Node Embeddings)概述，9页pdf

专知

15+阅读 · 2020年8月22日

【CVPR2020-旷视】DPGN：分布传播图网络的小样本学习

【CVPR2020-旷视】DPGN：分布传播图网络的小样本学习

专知

13+阅读 · 2020年4月1日

数据分析师应该知道的16种回归方法：负二项回归

数据分析师应该知道的16种回归方法：负二项回归

数萃大数据

74+阅读 · 2018年9月16日

基于径向基函数无网格离散的快速多水平算法

国家自然科学基金

0+阅读 · 2015年12月31日

Schr？dinger-Poisson方程守恒DDG方法研究

国家自然科学基金

2+阅读 · 2015年12月31日

光滑函数类的熵数估计

国家自然科学基金

0+阅读 · 2015年12月31日

一般误差分布下若干半参数模型的复合分位数方法

国家自然科学基金

0+阅读 · 2014年12月31日

Poisson流形上的修正Hamilton方法

国家自然科学基金

0+阅读 · 2014年12月31日

Learning Mixture Models via Efficient High-dimensional Sparse Fourier Transforms

Arxiv

0+阅读 · 1月8日

High-Dimensional Change Point Detection using Graph Spanning Ratio

Arxiv

0+阅读 · 1月8日

Local Interpolation via Low-Rank Tensor Trains

Arxiv

0+阅读 · 1月7日

High-Dimensional Precision Matrix Quadratic Forms: Estimation Framework for $p > n$

Arxiv

0+阅读 · 1月7日

Non-Homogeneous Markov-Switching Generalized Additive Models for Location, Scale, and Shape

Arxiv

0+阅读 · 1月7日

VIP会员

文章信息

相关主题

相关VIP内容

【ICML2025】免费的Fisher？通过回收平方梯度累加器近似Fisher信息矩阵

【ICML2025】免费的Fisher？通过回收平方梯度累加器近似Fisher信息矩阵

专知会员服务

12+阅读 · 2025年7月28日

UnHiPPO：面向不确定性的状态空间模型初始化方法

UnHiPPO：面向不确定性的状态空间模型初始化方法

专知会员服务

11+阅读 · 2025年6月6日

【NeurIPS2022】黎曼扩散模型

【NeurIPS2022】黎曼扩散模型

专知会员服务

43+阅读 · 2022年9月15日

NeurIPS 2021 | 寻找用于变分布泛化的隐式因果因子

NeurIPS 2021 | 寻找用于变分布泛化的隐式因果因子

专知会员服务

17+阅读 · 2021年12月7日

【ICML2021】随机傅立叶特征的量化算法

专知会员服务

25+阅读 · 2021年7月31日

热门VIP内容

开通专知VIP会员享更多权益服务

《面向小规模遥感应用引入思维链推理与多模态小语言模型》

《大国竞争时代的美国太空竞争力》50页报告

网络中心战：未来冲突

《自主无人机不会取代战斗机飞行员，将成为其僚机：协同作战飞机是下一代无人作战飞机》报告

相关资讯

【CVPR2021】CausalVAE: 引入因果结构的解耦表示学习

【CVPR2021】CausalVAE: 引入因果结构的解耦表示学习

专知

19+阅读 · 2021年3月28日

【CVPR2021】跨模态检索的概率嵌入

【CVPR2021】跨模态检索的概率嵌入

专知

17+阅读 · 2021年3月2日

图节点嵌入(Node Embeddings)概述，9页pdf

图节点嵌入(Node Embeddings)概述，9页pdf

专知

15+阅读 · 2020年8月22日

【CVPR2020-旷视】DPGN：分布传播图网络的小样本学习

【CVPR2020-旷视】DPGN：分布传播图网络的小样本学习

专知

13+阅读 · 2020年4月1日

数据分析师应该知道的16种回归方法：负二项回归

数据分析师应该知道的16种回归方法：负二项回归

数萃大数据

74+阅读 · 2018年9月16日

相关论文

Learning Mixture Models via Efficient High-dimensional Sparse Fourier Transforms

Arxiv

0+阅读 · 1月8日

High-Dimensional Change Point Detection using Graph Spanning Ratio

Arxiv

0+阅读 · 1月8日

Local Interpolation via Low-Rank Tensor Trains

Arxiv

0+阅读 · 1月7日

High-Dimensional Precision Matrix Quadratic Forms: Estimation Framework for $p > n$

Arxiv

0+阅读 · 1月7日

Non-Homogeneous Markov-Switching Generalized Additive Models for Location, Scale, and Shape

Arxiv

0+阅读 · 1月7日

相关基金

基于径向基函数无网格离散的快速多水平算法

国家自然科学基金

0+阅读 · 2015年12月31日

Schr？dinger-Poisson方程守恒DDG方法研究

国家自然科学基金

2+阅读 · 2015年12月31日

光滑函数类的熵数估计

国家自然科学基金

0+阅读 · 2015年12月31日

一般误差分布下若干半参数模型的复合分位数方法

国家自然科学基金

0+阅读 · 2014年12月31日

Poisson流形上的修正Hamilton方法

国家自然科学基金

0+阅读 · 2014年12月31日

微信扫码咨询专知VIP会员