Diverse Similarity Encoder for Deep GAN Inversion

Cheng Yu,Wenmin Wang,Roberto Bugiolacchi
DOI: https://doi.org/10.1016/j.asoc.2024.112201
2024-12-12
Abstract:Current deep generative adversarial networks (GANs) can synthesize high-quality (HQ) images, so learning representation with GANs is favorable. GAN inversion is one of emerging approaches that study how to invert images into latent space. Existing GAN encoders can invert images on StyleGAN, but cannot adapt to other deep GANs. We propose a novel approach to address this issue. By evaluating diverse similarity in latent vectors and images, we design an adaptive encoder, named diverse similarity encoder (DSE), that can be expanded to a variety of state-of-the-art GANs. DSE makes GANs reconstruct higher fidelity images from HQ images, no matter whether they are synthesized or real images. DSE has unified convolutional blocks and adapts well to mainstream deep GANs, e.g., PGGAN, StyleGAN, and BigGAN.
Computer Vision and Pattern Recognition,Image and Video Processing
What problem does this paper attempt to address?