Pith. sign in

REVIEW

AE-StyleGAN: Improved Training of Style-Based Auto-Encoders

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.08718 v1 pith:B4TRNBLF submitted 2021-10-17 cs.CV eess.IV

classification cs.CVeess.IV
keywords generatorlatentspacestyle-baseddatadisentangledencodergeneration
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

StyleGANs have shown impressive results on data generation and manipulation in recent years, thanks to its disentangled style latent space. A lot of efforts have been made in inverting a pretrained generator, where an encoder is trained ad hoc after the generator is trained in a two-stage fashion. In this paper, we focus on style-based generators asking a scientific question: Does forcing such a generator to reconstruct real data lead to more disentangled latent space and make the inversion process from image to latent space easy? We describe a new methodology to train a style-based autoencoder where the encoder and generator are optimized end-to-end. We show that our proposed model consistently outperforms baselines in terms of image inversion and generation quality. Supplementary, code, and pretrained models are available on the project website.

Discussion (0). Continue with ORCID to comment.

Pith tools