REVIEW 3 cited by
SinGAN: Learning a Generative Model from a Single Natural Image
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We introduce SinGAN, an unconditional generative model that can be learned from a single natural image. Our model is trained to capture the internal distribution of patches within the image, and is then able to generate high quality, diverse samples that carry the same visual content as the image. SinGAN contains a pyramid of fully convolutional GANs, each responsible for learning the patch distribution at a different scale of the image. This allows generating new samples of arbitrary size and aspect ratio, that have significant variability, yet maintain both the global structure and the fine textures of the training image. In contrast to previous single image GAN schemes, our approach is not limited to texture images, and is not conditional (i.e. it generates samples from noise). User studies confirm that the generated samples are commonly confused to be real images. We illustrate the utility of SinGAN in a wide range of image manipulation tasks.
Forward citations
Cited by 3 Pith papers
-
GarmentZoom: Generating Zoomable Images from Garment Listings
GarmentZoom trains one model to synthesize unaligned close-up details into full-view garment images across continuous scales 3-20x without per-instance tuning.
-
Leveraging Diffusion Models for Stylization using Multiple Style Images
A multi-image diffusion stylization pipeline that averages style embeddings, fine-tunes an IPAdapter, and clusters self-attention key/value features from style images achieves state-of-the-art scores on a new style-tr...
-
UltraZoom: Generating Gigapixel Images from Regular Photos
UltraZoom generates coherent gigapixel imagery from a regular full view and sparse close-ups by per-instance fine-tuning of a pretrained generative model with video-based registration.
Discussion (0). Sign in to comment.