Pith. sign in

REVIEW 1 cited by

S2F2: Self-Supervised High Fidelity Face Reconstruction from Monocular Image

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2203.07732 v2 pith:3F6NNP4Y submitted 2022-03-15 cs.CV cs.GRcs.LG

S2F2: Self-Supervised High Fidelity Face Reconstruction from Monocular Image

classification cs.CV cs.GRcs.LG
keywords facereconstructionimagefidelityhighself-supervisedgeometrymethod
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

We present a novel face reconstruction method capable of reconstructing detailed face geometry, spatially varying face reflectance from a single monocular image. We build our work upon the recent advances of DNN-based auto-encoders with differentiable ray tracing image formation, trained in self-supervised manner. While providing the advantage of learning-based approaches and real-time reconstruction, the latter methods lacked fidelity. In this work, we achieve, for the first time, high fidelity face reconstruction using self-supervised learning only. Our novel coarse-to-fine deep architecture allows us to solve the challenging problem of decoupling face reflectance from geometry using a single image, at high computational speed. Compared to state-of-the-art methods, our method achieves more visually appealing reconstruction.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. RelightAnyone: A Generalized Relightable 3D Gaussian Head Model

    cs.CV 2026-01 conditional novelty 6.0

    A two-stage model turns flat-lit photos of a new head into a relightable 3D Gaussian avatar, predicting reflectance parameters without needing one-light-at-a-time captures of that person.