Pith. sign in

REVIEW 2 cited by

Balancing Reconstruction Quality and Regularisation in ELBO for VAEs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1909.03765 v1 pith:KG6OFTNK submitted 2019-09-09 cs.LG cs.CVstat.ML

classification cs.LGcs.CVstat.ML
keywords variancenoisereconstructionelbolearningpriorqualitybalance
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

A trade-off exists between reconstruction quality and the prior regularisation in the Evidence Lower Bound (ELBO) loss that Variational Autoencoder (VAE) models use for learning. There are few satisfactory approaches to deal with a balance between the prior and reconstruction objective, with most methods dealing with this problem through heuristics. In this paper, we show that the noise variance (often set as a fixed value) in the Gaussian likelihood p(x|z) for real-valued data can naturally act to provide such a balance. By learning this noise variance so as to maximise the ELBO loss, we automatically obtain an optimal trade-off between the reconstruction error and the prior constraint on the posteriors. This variance can be interpreted intuitively as the necessary noise level for the current model to be the best explanation of the observed dataset. Further, by allowing the variance inference to be more flexible it can conveniently be used as an uncertainty estimator for reconstructed or generated samples. We demonstrate that optimising the noise variance is a crucial component of VAE learning, and showcase the performance on MNIST, Fashion MNIST and CelebA datasets. We find our approach can significantly improve the quality of generated samples whilst maintaining a smooth latent-space manifold to represent the data. The method also offers an indication of uncertainty in the final generative model.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Geometry-Preserving Encoder/Decoder in Latent Generative Models

    math.NA 2025-01 conditional novelty 6.0 of 10

    A geometry-preserving encoder/decoder trained by minimizing a logarithmic Gromov-Monge distance converges provably and reconstructs images faster than VAE-based latent generative models.

  2. Multipath Adaptive Gated Bottleneck Latent ODE with Raman Data Fusion for Cell Culture Process Forecasting

    cs.LG 2026-06 unverdicted novelty 5.5 of 10

    MP-JIT-FT with a gated-bottleneck Latent ODE and Raman fusion ranks best and beats a global Latent ODE on 8/9 targets across 38 heterogeneous fed-batch runs.

Pith tools