REVIEW 2 cited by
Animal Avatars: Reconstructing Animatable 3D Animals from Casual Videos
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We present a method to build animatable dog avatars from monocular videos. This is challenging as animals display a range of (unpredictable) non-rigid movements and have a variety of appearance details (e.g., fur, spots, tails). We develop an approach that links the video frames via a 4D solution that jointly solves for animal's pose variation, and its appearance (in a canonical pose). To this end, we significantly improve the quality of template-based shape fitting by endowing the SMAL parametric model with Continuous Surface Embeddings, which brings image-to-mesh reprojection constaints that are denser, and thus stronger, than the previously used sparse semantic keypoint correspondences. To model appearance, we propose an implicit duplex-mesh texture that is defined in the canonical pose, but can be deformed using SMAL pose coefficients and later rendered to enforce a photometric compatibility with the input video frames. On the challenging CoP3D and APTv2 datasets, we demonstrate superior results (both in terms of pose estimates and predicted appearance) to existing template-free (RAC) and template-based approaches (BARC, BITE).
Forward citations
Cited by 2 Pith papers
-
RatBodyFormer: Rat Body Surface from Keypoints
A new multi-camera dataset and a transformer-based method reconstruct a dense 3D rat body surface from 10 sparse keypoints, with reported mean errors around 5 to 7 mm.
-
AniMer: Animal Pose and Shape Estimation Using Family Aware Transformer
A family-aware Transformer with supervised contrastive learning and a diffusion-generated synthetic dataset achieves state-of-the-art 3D animal pose and shape estimation.
Discussion (0). Continue with ORCID to comment.