Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

DreamSteerer: Enhancing Source Image Conditioned Editability using Personalized Diffusion Models

About

Recent text-to-image personalization methods have shown great promise in teaching a diffusion model user-specified concepts given a few images for reusing the acquired concepts in a novel context. With massive efforts being dedicated to personalized generation, a promising extension is personalized editing, namely to edit an image using personalized concepts, which can provide a more precise guidance signal than traditional textual guidance. To address this, a straightforward solution is to incorporate a personalized diffusion model with a text-driven editing framework. However, such a solution often shows unsatisfactory editability on the source image. To address this, we propose DreamSteerer, a plug-in method for augmenting existing T2I personalization methods. Specifically, we enhance the source image conditioned editability of a personalized diffusion model via a novel Editability Driven Score Distillation (EDSD) objective. Moreover, we identify a mode trapping issue with EDSD, and propose a mode shifting regularization with spatial feature guided sampling to avoid such an issue. We further employ two key modifications to the Delta Denoising Score framework that enable high-fidelity local editing with personalized concepts. Extensive experiments validate that DreamSteerer can significantly improve the editability of several T2I personalization baselines while being computationally efficient.

Zhengyang Yu, Zhaoyuan Yang, Jing Zhang• 2024

Related benchmarks

TaskDatasetResultRank
Personalized Image EditingPersonalized Image Editing Dataset (test)
CLIP Score (B/32)79.6
6
Personalized Image EditingPersonalization One-shot scenario
CLIP-I (B/32)0.801
4
User preference studyTextual Inversion dataset (test)
User Preference Score3.78
2
User preference studyDreamBooth (test)
User Preference Score3.76
2
User preference studyCustom Diffusion dataset (test)
User Preference Score3.78
2
Showing 5 of 5 rows

Other info

Code

Follow for update