arrow
Return

Guided Safe Diffusion: Prohibiting Diffusion Models from Generating Inappropriate Content

delete2026-01-01
delete0
PRE
AI
S
Sidong Jiang
张瑞 cover
张瑞 (Rui Zhang)
X
Xi Yang
B
Bin Dong
K
Kaizhu Huang *
DOI:10.1007/978-981-96-7036-9_11delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The increasing deployment of large generative models has heightened concerns over security and privacy, particularly regarding the generation of inappropriate content such as violent, explicit, or sensitive images, as well as the potential for creating fake images that spread misinformation and cause social problems. In this work, we propose Guided Safe Diffusion (GSD), an inference-time method specifically designed for diffusion models to prevent the generation of images with undesirable content as defined in a prohibited content list. Our method integrates safety guidance during the denoising steps of the model's inference process, modifying the predicted noise to steer the generation process away from unwanted content. This approach allows the model to accept both an input image and a text description, facilitating controlled image generation. Unlike previous methods, our technique does not necessitate retraining or fine-tuning of the model. We conduct qualitative and quantitative experiments to assess the effectiveness of our method, demonstrating that GSD can remove the unwanted content while preserving unrelated content. The results validate our method's ability to mitigate risks while maintaining the generative utility of diffusion models.
Keywords:
Image protection
Diffusion models
AI ethics

Journal

N
NEURAL INFORMATION PROCESSING, ICONIP 2024, PT XVI
IF:
0
Papers:
20
Citations:
0

Organization

X
xi'an jiaotong-liverpool university
Scholars:
965
Papers: 515
Citations: 0
D
Duke Kunshan University
Scholars:
1.1K
Papers: 982
Citations: 1.5K