A network that feeds three pooled multi-scale CNN features through a two-block transformer reaches PC 0.9187 on SCUT-FBP5500, edging out the cited R3CNN baseline.
Facial Beauty Prediction Using Global Context Vision Transformer,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Scale-interaction transformer: a hybrid cnn-transformer model for facial beauty prediction
A network that feeds three pooled multi-scale CNN features through a two-block transformer reaches PC 0.9187 on SCUT-FBP5500, edging out the cited R3CNN baseline.