Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation

About

We present a generic image-to-image translation framework, pixel2style2pixel (pSp). Our pSp framework is based on a novel encoder network that directly generates a series of style vectors which are fed into a pretrained StyleGAN generator, forming the extended W+ latent space. We first show that our encoder can directly embed real images into W+, with no additional optimization. Next, we propose utilizing our encoder to directly solve image-to-image translation tasks, defining them as encoding problems from some input domain into the latent domain. By deviating from the standard invert first, edit later methodology used with previous StyleGAN encoders, our approach can handle a variety of tasks even when the input image is not represented in the StyleGAN domain. We show that solving translation tasks through StyleGAN significantly simplifies the training process, as no adversary is required, has better support for solving tasks without pixel-to-pixel correspondence, and inherently supports multi-modal synthesis via the resampling of styles. Finally, we demonstrate the potential of our framework on a variety of facial image-to-image translation tasks, even when compared to state-of-the-art solutions designed specifically for a single task, and further show that it can be extended beyond the human facial domain.

Elad Richardson, Yuval Alaluf, Or Patashnik, Yotam Nitzan, Yaniv Azar, Stav Shapiro, Daniel Cohen-Or• 2020

Related benchmarks

TaskDatasetResultRank
Image ReconstructionFFHQ No glasses
LPIPS0.147
18
Image ReconstructionFFHQ Glasses
LPIPS0.15
18
Attribute ClassificationFFHQ (test)
Accuracy85
15
Image Editing (Remove glasses)FFHQ (test)
ID-Sim0.625
15
Image Editing (Add glasses)FFHQ (test)
ID-Sim0.418
15
Face image reconstructionCelebA-HQ (test)
MAE0.079
13
Real image projectionCelebA-HQ (test)
MSE0.037
9
Sketch-to-Photo GenerationChair V2
FID105.5
8
Sketch-to-Photo GenerationShoe V2
FID54.48
8
Sketch-to-Photo GenerationHandbag
FID122.5
8
Showing 10 of 24 rows

Other info

Follow for update