REVIEW 3 cited by
Generalizable Neural Voxels for Fast Human Radiance Fields
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Rendering moving human bodies at free viewpoints only from a monocular video is quite a challenging problem. The information is too sparse to model complicated human body structures and motions from both view and pose dimensions. Neural radiance fields (NeRF) have shown great power in novel view synthesis and have been applied to human body rendering. However, most current NeRF-based methods bear huge costs for both training and rendering, which impedes the wide applications in real-life scenarios. In this paper, we propose a rendering framework that can learn moving human body structures extremely quickly from a monocular video. The framework is built by integrating both neural fields and neural voxels. Especially, a set of generalizable neural voxels are constructed. With pretrained on various human bodies, these general voxels represent a basic skeleton and can provide strong geometric priors. For the fine-tuning process, individual voxels are constructed for learning differential textures, complementary to general voxels. Thus learning a novel body can be further accelerated, taking only a few minutes. Our method shows significantly higher training efficiency compared with previous methods, while maintaining similar rendering quality. The project page is at https://taoranyi.com/gneuvox .
Forward citations
Cited by 3 Pith papers
-
Style4D-Bench: A Benchmark Suite for 4D Stylization
Style4D-Bench introduces a 12-metric evaluation protocol and a 4DGS-based baseline, Style4D, claimed to achieve state-of-the-art 4D stylization.
-
EPSilon: Efficient Point Sampling for Lightening of Hybrid-based 3D Avatar Generation
EPSilon prunes empty rays and sampling intervals around the body mesh, cutting hybrid avatar rendering to 3.9% of the points and 20x faster inference with comparable quality.
-
Snap-Snap: Taking Two Images to Reconstruct 3D Human Gaussians in Milliseconds
A feed-forward pipeline predicts 3D human Gaussian splats from two input images (front and back) in 190 ms, using a DUSt3R-style point cloud predictor with extra side-view heads, nearest-neighbor color warping, and a ...
Discussion (0). Sign in to comment.