Pith. sign in

REVIEW 2 cited by

GAN Inversion: A Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2101.05278 v5 pith:NHYY2BV3 submitted 2021-01-14 cs.CV

classification cs.CV
keywords imageinversionapplicationslatentpretrainedrealspaceaims
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

GAN inversion aims to invert a given image back into the latent space of a pretrained GAN model, for the image to be faithfully reconstructed from the inverted code by the generator. As an emerging technique to bridge the real and fake image domains, GAN inversion plays an essential role in enabling the pretrained GAN models such as StyleGAN and BigGAN to be used for real image editing applications. Meanwhile, GAN inversion also provides insights on the interpretation of GAN's latent space and how the realistic images can be generated. In this paper, we provide an overview of GAN inversion with a focus on its recent algorithms and applications. We cover important techniques of GAN inversion and their applications to image restoration and image manipulation. We further elaborate on some trends and challenges for future directions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A framework for river connectivity classification using temporal image processing and attention based neural networks

    cs.CV 2025-02 conditional novelty 5.0 of 10

    A trail camera image pipeline using temporal luma enhancement and a vision transformer classifies unseen river sites as connected or disconnected with about 90% reported accuracy.

  2. Hands-off Image Editing: Language-guided Editing without any Task-specific Labeling, Masking or even Training

    cs.CL 2025-02 conditional novelty 4.0 of 10

    An instruction-guided image editor that needs no training, labels, or masks: an LLM writes before/after captions and their embedding difference guides Stable Diffusion.

Pith tools