REVIEW 2 cited by
Animatable Implicit Neural Representations for Creating Realistic Avatars from Videos
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper addresses the challenge of reconstructing an animatable human model from a multi-view video. Some recent works have proposed to decompose a non-rigidly deforming scene into a canonical neural radiance field and a set of deformation fields that map observation-space points to the canonical space, thereby enabling them to learn the dynamic scene from images. However, they represent the deformation field as translational vector field or SE(3) field, which makes the optimization highly under-constrained. Moreover, these representations cannot be explicitly controlled by input motions. Instead, we introduce a pose-driven deformation field based on the linear blend skinning algorithm, which combines the blend weight field and the 3D human skeleton to produce observation-to-canonical correspondences. Since 3D human skeletons are more observable, they can regularize the learning of the deformation field. Moreover, the pose-driven deformation field can be controlled by input skeletal motions to generate new deformation fields to animate the canonical human model. Experiments show that our approach significantly outperforms recent human modeling methods. The code is available at https://zju3dv.github.io/animatable_nerf/.
Forward citations
Cited by 2 Pith papers
-
Real-Time Human Reconstruction and Animation using Feed-Forward Gaussian Splatting
A feed-forward transformer predicts SMPL-X vertex-aligned 3D Gaussians in a canonical T-pose, enabling real-time animation by linear blend skinning without per-frame network inference.
-
Deblur-Avatar: Animatable Avatars from Motion-Blurred Monocular Videos
Deblur-Avatar reconstructs sharp, animatable human avatars from motion-blurred monocular video by optimizing SMPL start and end poses and averaging rendered virtual frames inside 3D Gaussian Splatting.
Discussion (0). Continue with ORCID to comment.