Pith. sign in

REVIEW 1 cited by

Vid2Actor: Free-viewpoint Animatable Person Synthesis from Video in the Wild

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2012.12884 v1 pith:VSWK6D3A submitted 2020-12-23 cs.CV cs.GR

classification cs.CVcs.GR
keywords videomodelpersonposesynthesisanimatablecameralearned
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Given an "in-the-wild" video of a person, we reconstruct an animatable model of the person in the video. The output model can be rendered in any body pose to any camera view, via the learned controls, without explicit 3D mesh reconstruction. At the core of our method is a volumetric 3D human representation reconstructed with a deep network trained on input video, enabling novel pose/view synthesis. Our method is an advance over GAN-based image-to-image translation since it allows image synthesis for any pose and camera via the internal 3D representation, while at the same time it does not require a pre-rigged model or ground truth meshes for training, as in mesh-based learning. Experiments validate the design choices and yield results on synthetic data and on real videos of diverse people performing unconstrained activities (e.g. dancing or playing tennis). Finally, we demonstrate motion re-targeting and bullet-time rendering with the learned models.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SkinningGS: Editable Dynamic Human Scene Reconstruction Using Gaussian Splatting Based on a Skinning Model

    cs.GR 2025-06 conditional novelty 5.0 of 10

    A UV-texture-driven Gaussian splatting avatar method claims faster, leaner, and better human-scene reconstruction than HUGS, but its tables contain internal inconsistencies.

Pith tools