Pith. sign in

REVIEW 1 cited by

Frame attention networks for facial expression recognition in videos

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1907.00193 v2 pith:TVBZFBY6 submitted 2019-06-29 cs.CV cs.HCcs.MM

Frame attention networks for facial expression recognition in videos

classification cs.CV cs.HCcs.MM
keywords attentionfacialfeatureframenetworkvideodiscriminativeexpression
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The video-based facial expression recognition aims to classify a given video into several basic emotions. How to integrate facial features of individual frames is crucial for this task. In this paper, we propose the Frame Attention Networks (FAN), to automatically highlight some discriminative frames in an end-to-end framework. The network takes a video with a variable number of face images as its input and produces a fixed-dimension representation. The whole network is composed of two modules. The feature embedding module is a deep Convolutional Neural Network (CNN) which embeds face images into feature vectors. The frame attention module learns multiple attention weights which are used to adaptively aggregate the feature vectors to form a single discriminative video representation. We conduct extensive experiments on CK+ and AFEW8.0 datasets. Our proposed FAN shows superior performance compared to other CNN based methods and achieves state-of-the-art performance on CK+.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Bootstrap Model Ensemble and Rank Loss for Engagement Intensity Regression

    cs.CV 2019-07 unverdicted novelty 3.0

    The authors achieve third place in EmotiW 2019 engagement intensity regression by extending an LSTM framework with facial landmarks, rank loss, and bootstrap aggregation to reach MSE 0.0626 on the test set.