Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation

Richardson, Elad; Alaluf, Yuval; Patashnik, Or; Nitzan, Yotam; Azar, Yaniv; Shapiro, Stav; Cohen-Or, Daniel

Computer Science > Computer Vision and Pattern Recognition

arXiv:2008.00951v1 (cs)

[Submitted on 3 Aug 2020 (this version), latest version 21 Apr 2021 (v2)]

Title:Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation

Authors:Elad Richardson, Yuval Alaluf, Or Patashnik, Yotam Nitzan, Yaniv Azar, Stav Shapiro, Daniel Cohen-Or

View PDF

Abstract:We present a generic image-to-image translation framework, Pixel2Style2Pixel (pSp). Our pSp framework is based on a novel encoder network that directly generates a series of style vectors which are fed into a pretrained StyleGAN generator, forming the extended W+ latent space. We first show that our encoder can directly embed real images into W+, with no additional optimization. We further introduce a dedicated identity loss which is shown to achieve improved performance in the reconstruction of an input image. We demonstrate pSp to be a simple architecture that, by leveraging a well-trained, fixed generator network, can be easily applied on a wide-range of image-to-image translation tasks. Solving these tasks through the style representation results in a global approach that does not rely on a local pixel-to-pixel correspondence and further supports multi-modal synthesis via the resampling of styles. Notably, we demonstrate that pSp can be trained to align a face image to a frontal pose without any labeled data, generate multi-modal results for ambiguous tasks such as conditional face generation from segmentation maps, and construct high-resolution images from corresponding low-resolution images.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2008.00951 [cs.CV]
	(or arXiv:2008.00951v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2008.00951

Submission history

From: Elad Richardson [view email]
[v1] Mon, 3 Aug 2020 15:30:38 UTC (13,561 KB)
[v2] Wed, 21 Apr 2021 12:53:36 UTC (33,056 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators