REVIEW 4 cited by
GaussianHead: High-fidelity Head Avatars with Learnable Gaussian Derivation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Constructing vivid 3D head avatars for given subjects and realizing a series of animations on them is valuable yet challenging. This paper presents GaussianHead, which models the actional human head with anisotropic 3D Gaussians. In our framework, a motion deformation field and multi-resolution tri-plane are constructed respectively to deal with the head's dynamic geometry and complex texture. Notably, we impose an exclusive derivation scheme on each Gaussian, which generates its multiple doppelgangers through a set of learnable parameters for position transformation. With this design, we can compactly and accurately encode the appearance information of Gaussians, even those fitting the head's particular components with sophisticated structures. In addition, an inherited derivation strategy for newly added Gaussians is adopted to facilitate training acceleration. Extensive experiments show that our method can produce high-fidelity renderings, outperforming state-of-the-art approaches in reconstruction, cross-identity reenactment, and novel view synthesis tasks. Our code is available at: https://github.com/chiehwangs/gaussian-head.
Forward citations
Cited by 4 Pith papers
-
Detangled: A Framework for Creating, Editing, and Inferencing Feature Rich Hair Strands
A 5D texture parameterization plus centerline-based canonical space and supervised diffusion enables generation and texture transfer of feature-rich hair strands independent of style.
-
GeoAvatar: Adaptive Geometrical Gaussian Splatting for 3D Head Avatar
GeoAvatar improves 3D head avatar quality by adaptively regulating Gaussian offsets per facial region, adding a detailed mouth structure with part-wise deformation, and releasing a new expressive monocular dataset, Dy...
-
PD-GS: Phoneme-Driven 3DGS for Audio-Driven Talking Heads
PD-GS fuses ASR-aligned phoneme tokens with HuBERT audio features via a learned gate in a 3DGS talker, improving lip landmark distance (LMD 2.66 on HDTF) over audio-only baselines.
-
EVA: Expressive Virtual Avatars from Multi-view Videos
EVA reconstructs an actor-specific avatar from multi-view video that renders in real time and allows independent control of body motion, hand gestures, and facial expressions through a disentangled two-layer Gaussian ...
Discussion (0). Continue with ORCID to comment.