Pith. sign in

REVIEW 2 cited by

Subpixel Heatmap Regression for Facial Landmark Localization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2111.02360 v1 pith:VCEHYD7S submitted 2021-11-03 cs.CV

classification cs.CV
keywords heatmapfaciallandmarklocalizationmodelsregressionacrossapproach
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep Learning models based on heatmap regression have revolutionized the task of facial landmark localization with existing models working robustly under large poses, non-uniform illumination and shadows, occlusions and self-occlusions, low resolution and blur. However, despite their wide adoption, heatmap regression approaches suffer from discretization-induced errors related to both the heatmap encoding and decoding process. In this work we show that these errors have a surprisingly large negative impact on facial alignment accuracy. To alleviate this problem, we propose a new approach for the heatmap encoding and decoding process by leveraging the underlying continuous distribution. To take full advantage of the newly proposed encoding-decoding mechanism, we also introduce a Siamese-based training that enforces heatmap consistency across various geometric image transformations. Our approach offers noticeable gains across multiple datasets setting a new state-of-the-art result in facial landmark localization. Code alongside the pretrained models will be made available at https://www.adrianbulat.com/face-alignment

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ORFormer: Occlusion-Robust Transformer for Accurate Facial Landmark Detection

    cs.CV 2024-12 conditional novelty 6.0 of 10

    ORFormer uses per-patch messenger tokens to detect and recover occluded face regions, reducing landmark error on WFLW and COFW.

  2. landmarker: a Toolkit for Anatomical Landmark Localization in 2D/3D Images

    cs.CV 2025-01 conditional novelty 5.0 of 10

    landmarker provides a modular PyTorch-based toolkit for anatomical landmark localization in 2D/3D medical images, and its included models outperform literature baselines on two benchmark datasets.

Pith tools