Pith. sign in

REVIEW 1 cited by

FaceDancer: Pose- and Occlusion-Aware High Fidelity Face Swapping

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.10473 v2 pith:6IGFSWQJ submitted 2022-10-19 cs.CV

classification cs.CV
keywords faceidentityfacedancerfeaturesaffafacialfeaturefidelity
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this work, we present a new single-stage method for subject agnostic face swapping and identity transfer, named FaceDancer. We have two major contributions: Adaptive Feature Fusion Attention (AFFA) and Interpreted Feature Similarity Regularization (IFSR). The AFFA module is embedded in the decoder and adaptively learns to fuse attribute features and features conditioned on identity information without requiring any additional facial segmentation process. In IFSR, we leverage the intermediate features in an identity encoder to preserve important attributes such as head pose, facial expression, lighting, and occlusion in the target face, while still transferring the identity of the source face with high fidelity. We conduct extensive quantitative and qualitative experiments on various datasets and show that the proposed FaceDancer outperforms other state-of-the-art networks in terms of identityn transfer, while having significantly better pose preservation than most of the previous methods.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. DFCon: Attention-Driven Supervised Contrastive Learning for Robust Deepfake Detection

    cs.CV 2025-01 conditional novelty 3.0 of 10

    An ensemble of three pretrained vision transformers trained with supervised contrastive loss and majority voting reports 95.83% validation accuracy on the DFWild-Cup 2025 deepfake detection dataset.

Pith tools