REVIEW 3 cited by
Ultraman: Single Image 3D Human Reconstruction with Ultra Speed and Detail
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
3D human body reconstruction has been a challenge in the field of computer vision. Previous methods are often time-consuming and difficult to capture the detailed appearance of the human body. In this paper, we propose a new method called \emph{Ultraman} for fast reconstruction of textured 3D human models from a single image. Compared to existing techniques, \emph{Ultraman} greatly improves the reconstruction speed and accuracy while preserving high-quality texture details. We present a set of new frameworks for human reconstruction consisting of three parts, geometric reconstruction, texture generation and texture mapping. Firstly, a mesh reconstruction framework is used, which accurately extracts 3D human shapes from a single image. At the same time, we propose a method to generate a multi-view consistent image of the human body based on a single image. This is finally combined with a novel texture mapping method to optimize texture details and ensure color consistency during reconstruction. Through extensive experiments and evaluations, we demonstrate the superior performance of \emph{Ultraman} on various standard datasets. In addition, \emph{Ultraman} outperforms state-of-the-art methods in terms of human rendering quality and speed. Upon acceptance of the article, we will make the code and data publicly available.
Forward citations
Cited by 3 Pith papers
-
Diffusion-based Visual Anagram as Multi-task Learning
A diffusion-based method generates visual anagrams by treating each viewpoint as a task and adding anti-segregation, noise-balancing, and variance-rectification steps.
-
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
AniGS produces an animatable 3D avatar from a single image by synthesizing multi-view canonical images and normals with a video diffusion model and reconstructing them via 4D Gaussian Splatting.
-
DRiVE: Diffusion-based Rigging Empowers Generation of Versatile and Expressive Characters
A diffusion-based pipeline that rigs 3D Gaussian characters, including hair and clothing, using a newly curated dataset of 9,420 anime meshes.
Discussion (0). Continue with ORCID to comment.