Spatial Latent Representations in Generative Adversarial Networks for Image Generation

03/25/2023
by   Maciej Sypetkowski, et al.
0

In the majority of GAN architectures, the latent space is defined as a set of vectors of given dimensionality. Such representations are not easily interpretable and do not capture spatial information of image content directly. In this work, we define a family of spatial latent spaces for StyleGAN2, capable of capturing more details and representing images that are out-of-sample in terms of the number and arrangement of object parts, such as an image of multiple faces or a face with more than two eyes. We propose a method for encoding images into our spaces, together with an attribute model capable of performing attribute editing in these spaces. We show that our spaces are effective for image manipulation and encode semantic information well. Our approach can be used on pre-trained generator models, and attribute edition can be done using pre-generated direction vectors making the barrier to entry for experimentation and use extremely low. We propose a regularization method for optimizing latent representations, which equalizes distributions of parts of latent spaces, making representations much closer to generated ones. We use it for encoding images into spatial spaces to obtain significant improvement in quality while keeping semantics and ability to use our attribute model for edition purposes. In total, using our methods gives encoding quality boost even as high as 30 methods, while keeping semantics. Additionally, we propose a StyleGAN2 training procedure on our spatial latent spaces, together with a custom spatial latent representation distribution to make spatially closer elements in the representation more dependent on each other than farther elements. Such approach improves the FID score by 29 consistent images of arbitrary sizes on spatially homogeneous datasets, like satellite imagery.

READ FULL TEXT

page 18

page 19

page 25

page 31

page 33

page 34

page 35

page 36

research
04/10/2018

RSGAN: Face Swapping and Editing using Face and Hair Representation in Latent Spaces

In this paper, we present an integrated system for automatically generat...
research
08/07/2022

Hierarchical Semantic Regularization of Latent Spaces in StyleGANs

Progress in GANs has enabled the generation of high-resolution photoreal...
research
07/03/2019

Semi-supervised Image Attribute Editing using Generative Adversarial Networks

Image attribute editing is a challenging problem that has been recently ...
research
05/07/2019

Spatially Constrained Generative Adversarial Networks for Conditional Image Generation

Image generation has raised tremendous attention in both academic and in...
research
11/16/2021

Delta-GAN-Encoder: Encoding Semantic Changes for Explicit Image Editing, using Few Synthetic Samples

Understating and controlling generative models' latent space is a comple...
research
08/03/2021

Toward Spatially Unbiased Generative Models

Recent image generation models show remarkable generation performance. H...
research
06/03/2019

DualDis: Dual-Branch Disentangling with Adversarial Learning

In computer vision, disentangling techniques aim at improving latent rep...

Please sign up or login with your details

Forgot password? Click here to reset