Pith. sign in

REVIEW 2 cited by

Semi-Supervised StyleGAN for Disentanglement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2003.03461 v3 pith:YSIKFUDG submitted 2020-03-06 cs.CV cs.LG

classification cs.CVcs.LG
keywords disentanglementlearningdisentangledhigh-resolutioncontrollablecrucialdatasetsgeneration
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Disentanglement learning is crucial for obtaining disentangled representations and controllable generation. Current disentanglement methods face several inherent limitations: difficulty with high-resolution images, primarily focusing on learning disentangled representations, and non-identifiability due to the unsupervised setting. To alleviate these limitations, we design new architectures and loss functions based on StyleGAN (Karras et al., 2019), for semi-supervised high-resolution disentanglement learning. We create two complex high-resolution synthetic datasets for systematic testing. We investigate the impact of limited supervision and find that using only 0.25%~2.5% of labeled data is sufficient for good disentanglement on both synthetic and real datasets. We propose new metrics to quantify generator controllability, and observe there may exist a crucial trade-off between disentangled representation learning and controllable generation. We also consider semantic fine-grained image editing to achieve better generalization to unseen images.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Symbolic Disentangled Representations for Images

    cs.CV 2024-12 conditional novelty 6.0 of 10

    ArSyD learns image representations where each generative factor is a separate hypervector, enabling property editing by vector exchange and dimension-agnostic disentanglement evaluation.

  2. CtrlNeRF: The Generative Neural Radiation Fields for the Controllable Synthesis of High-fidelity 3D-Aware Images

    cs.CV 2024-12 conditional novelty 4.0 of 10

    CtrlNeRF learns a single shared neural radiance field generator that can synthesize controllable, 3D-consistent images of multiple object classes and colors using label-embedded latent codes.

Pith tools