A new multi-camera dataset and a transformer-based method reconstruct a dense 3D rat body surface from 10 sparse keypoints, with reported mean errors around 5 to 7 mm.
Animal Avatars: Reconstructing Animatable 3D Animals from Casual Videos
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
We present a method to build animatable dog avatars from monocular videos. This is challenging as animals display a range of (unpredictable) non-rigid movements and have a variety of appearance details (e.g., fur, spots, tails). We develop an approach that links the video frames via a 4D solution that jointly solves for animal's pose variation, and its appearance (in a canonical pose). To this end, we significantly improve the quality of template-based shape fitting by endowing the SMAL parametric model with Continuous Surface Embeddings, which brings image-to-mesh reprojection constaints that are denser, and thus stronger, than the previously used sparse semantic keypoint correspondences. To model appearance, we propose an implicit duplex-mesh texture that is defined in the canonical pose, but can be deformed using SMAL pose coefficients and later rendered to enforce a photometric compatibility with the input video frames. On the challenging CoP3D and APTv2 datasets, we demonstrate superior results (both in terms of pose estimates and predicted appearance) to existing template-free (RAC) and template-based approaches (BARC, BITE).
citation-role summary
citation-polarity summary
fields
cs.CV 1years
2024 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
RatBodyFormer: Rat Body Surface from Keypoints
A new multi-camera dataset and a transformer-based method reconstruct a dense 3D rat body surface from 10 sparse keypoints, with reported mean errors around 5 to 7 mm.