Pith. sign in

REVIEW 2 cited by

Unsupervised Discovery of Interpretable Directions in the GAN Latent Space

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2002.03754 v3 pith:EXGSZZT5 submitted 2020-02-10 cs.LG cs.CVstat.ML

Unsupervised Discovery of Interpretable Directions in the GAN Latent Space

classification cs.LG cs.CVstat.ML
keywords directionslatentcorrespondingdiscoveryexistingforminterpretablemodels
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The latent spaces of GAN models often have semantically meaningful directions. Moving in these directions corresponds to human-interpretable image transformations, such as zooming or recoloring, enabling a more controllable generation process. However, the discovery of such directions is currently performed in a supervised manner, requiring human labels, pretrained models, or some form of self-supervision. These requirements severely restrict a range of directions existing approaches can discover. In this paper, we introduce an unsupervised method to identify interpretable directions in the latent space of a pretrained GAN model. By a simple model-agnostic procedure, we find directions corresponding to sensible semantic manipulations without any form of (self-)supervision. Furthermore, we reveal several non-trivial findings, which would be difficult to obtain by existing methods, e.g., a direction corresponding to background removal. As an immediate practical benefit of our work, we show how to exploit this finding to achieve competitive performance for weakly-supervised saliency detection.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Semantic Browsing: Controllable Diversity for Image Generation

    cs.CV 2026-06 unverdicted novelty 7.0

    A technique for controllable diversity in text-to-image generation by inducing structured semantic variations at the prompt level via VLM and agentic workflow.

  2. XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision

    cs.LG 2026-01 conditional novelty 6.0

    XFACTORS separates latent factors into per-factor subspaces with InfoNCE supervision, achieving near-perfect FactorVAE scores on synthetic benchmarks and qualitative factor swapping on CelebA.