FairGen: Enhancing Fairness in Text-to-Image Diffusion Models via Self-Discovering Latent Directions

While Diffusion Models (DM) exhibit remarkable performance across various image generative tasks, they nonetheless reflect the inherent bias presented in the training set. As DMs are now widely used in real-world applications, these biases could perpetuate a distorted worldview and hinder opportunities for minority groups. Existing methods on debiasing DMs usually requires model retraining with a human-crafted reference dataset or additional classifiers, which suffer from two major limitations: (1) collecting reference datasets causes expensive annotation cost; (2) the debiasing performance is heavily constrained by the quality of the reference dataset or the additional classifier. To address the above limitations, we propose FairGen, a plug-and-play method that learns attribute latent directions in a self-discovering manner, thus eliminating the reliance on such reference dataset. Specifically, FairGen consists of two parts: a set of attribute adapters and a distribution indicator. Each adapter in the set aims to learn an attribute latent direction, and is optimized via noise composition through a self-discovering process. Then, the distribution indicator is multiplied by the set of adapters to guide the generation process towards the prescribed distribution. Our method enables debiasing multiple attributes in DMs simultaneously, while remaining lightweight and easily integrable with other DMs, eliminating the need for retraining. Extensive experiments on debiasing gender, racial, and their intersectional biases show that our method outperforms previous SOTA by a large margin.

翻译：尽管扩散模型在各种图像生成任务中展现出卓越性能，但其仍会反映训练数据中存在的固有偏见。随着扩散模型在现实应用中的广泛使用，这些偏见可能延续扭曲的世界观并阻碍少数群体的发展机会。现有的扩散模型去偏方法通常需要基于人工构建的参考数据集或额外分类器进行模型重训练，这些方法存在两大局限：(1) 收集参考数据集会产生高昂的标注成本；(2) 去偏效果严重受限于参考数据集或额外分类器的质量。为克服上述局限，我们提出FairGen——一种通过自发现方式学习属性潜在方向的即插即用方法，从而消除对参考数据集的依赖。具体而言，FairGen由两部分组成：一组属性适配器和一个分布指示器。每个适配器旨在通过自发现过程中的噪声组合优化来学习特定属性的潜在方向。随后，分布指示器与适配器组相乘，引导生成过程朝向预设分布。本方法能够同时对扩散模型中的多重属性进行去偏处理，同时保持轻量化设计并易于与其他扩散模型集成，无需重新训练。在性别、种族及其交叉偏见的去偏实验中，大量实验表明我们的方法以显著优势超越先前的最先进技术。

相关内容

MoDELS

关注 45

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

【CVPR 2022】基于元内存传输的跨域少镜头语义分割，Remember the Difference: Cross-Domain Few-Shot Semantic Segmentation via Meta-Memory Transfer

专知会员服务

13+阅读 · 2022年3月12日

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

FlowQA: Grasping Flow in History for Conversational Machine Comprehension

专知会员服务

34+阅读 · 2019年10月18日