A video diffusion model fine-tuned to output both color and normal maps, aligned by a geometry-temporal attention block, reconstructs textured 3D meshes from a single image.
Zero-1-to-3: Zero-shot one image to 3d object,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
NOVA3D: Normal Aligned Video Diffusion Model for Single Image to 3D Generation
A video diffusion model fine-tuned to output both color and normal maps, aligned by a geometry-temporal attention block, reconstructs textured 3D meshes from a single image.