Image editing using diffusion models has witnessed extremely fast-paced growth recently. There are various ways in which previous works enable controlling and editing images. Some works use high-level conditioning such as text, while others use low-level conditioning. Nevertheless, most of them lack fine-grained control over the properties of the different objects present in the image, i.e. object-level image editing. In this work, we consider an image as a composition of multiple objects, each defined by various properties. Out of these properties, we identify structure and appearance as the most intuitive to understand and useful for editing purposes. We propose Structure-and-Appearance Paired Diffusion model (PAIR-Diffusion), which is trained using structure and appearance information explicitly extracted from the images. The proposed model enables users to inject a reference image's appearance into the input image at both the object and global levels. Additionally, PAIR-Diffusion allows editing the structure while maintaining the style of individual components of the image unchanged. We extensively evaluate our method on LSUN datasets and the CelebA-HQ face dataset, and we demonstrate fine-grained control over both structure and appearance at the object level. We also applied the method to Stable Diffusion to edit any real image at the object level.


翻译:近年来,利用扩散模型进行图像编辑的技术发展极为迅速。以往研究提供了多种控制与编辑图像的方式,部分工作采用文本等高层条件约束,另一部分则利用低层条件。然而,大多数方法缺乏对图像中不同物体属性的细粒度控制,即无法实现物体级图像编辑。本文将图像视为多个物体的组合,每个物体由不同属性定义。在这些属性中,我们识别出结构与外观是最直观且最有利于编辑的属性。我们提出结构与外观配对扩散模型(PAIR-Diffusion),该模型利用从图像中显式提取的结构与外观信息进行训练。所提模型支持用户将参考图像的外观注入输入图像中,既能实现物体级注入,也能实现全局级注入。此外,PAIR-Diffusion允许在保持图像中各组件风格不变的前提下编辑结构。我们在LSUN数据集和CelebA-HQ人脸数据集上对本方法进行了广泛评估,展示了在物体级对结构与外观的细粒度控制能力。我们还将该方法应用于Stable Diffusion,实现了对任意真实图像的物体级编辑。

0
下载
关闭预览

相关内容

两人亲密社交应用,官网: trypair.com/
【AAAI2023】用于复杂场景图像合成的特征金字塔扩散模型
专知会员服务
19+阅读 · 2021年9月13日
Hierarchically Structured Meta-learning
CreateAMind
27+阅读 · 2019年5月22日
Transferring Knowledge across Learning Processes
CreateAMind
29+阅读 · 2019年5月18日
Unsupervised Learning via Meta-Learning
CreateAMind
44+阅读 · 2019年1月3日
【推荐】深度学习目标检测全面综述
机器学习研究会
21+阅读 · 2017年9月13日
【推荐】全卷积语义分割综述
机器学习研究会
19+阅读 · 2017年8月31日
Generative Adversarial Text to Image Synthesis论文解读
统计学习与视觉计算组
13+阅读 · 2017年6月9日
国家自然科学基金
1+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
2+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2011年12月31日
国家自然科学基金
0+阅读 · 2009年12月31日
Arxiv
0+阅读 · 2023年5月18日
Arxiv
0+阅读 · 2023年5月18日
Arxiv
0+阅读 · 2023年5月17日
Arxiv
46+阅读 · 2022年9月6日
Arxiv
14+阅读 · 2022年8月25日
Arxiv
11+阅读 · 2018年5月13日
VIP会员
最新内容
《美军水下战与海床战概述及本地实施》
专知会员服务
0+阅读 · 42分钟前
面向未来冲突推进陆军情报体制改革
专知会员服务
0+阅读 · 今天4:12
乌克兰纵深打击如何重塑俄罗斯的战略选择
专知会员服务
2+阅读 · 7月24日
俄乌战争中关于中程打击无人机部署的经验启示
《基于强化学习的自动化红队测试》
专知会员服务
4+阅读 · 7月23日
伊朗不对称防空战略的演进
专知会员服务
4+阅读 · 7月23日
相关VIP内容
【AAAI2023】用于复杂场景图像合成的特征金字塔扩散模型
专知会员服务
19+阅读 · 2021年9月13日
相关基金
国家自然科学基金
1+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2013年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2012年12月31日
国家自然科学基金
2+阅读 · 2012年12月31日
国家自然科学基金
0+阅读 · 2011年12月31日
国家自然科学基金
0+阅读 · 2009年12月31日
Top
微信扫码咨询专知VIP会员