REVIEW 2 cited by
A Deep Learning Perspective on the Origin of Facial Expressions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Facial expressions play a significant role in human communication and behavior. Psychologists have long studied the relationship between facial expressions and emotions. Paul Ekman et al., devised the Facial Action Coding System (FACS) to taxonomize human facial expressions and model their behavior. The ability to recognize facial expressions automatically, enables novel applications in fields like human-computer interaction, social gaming, and psychological research. There has been a tremendously active research in this field, with several recent papers utilizing convolutional neural networks (CNN) for feature extraction and inference. In this paper, we employ CNN understanding methods to study the relation between the features these computational networks are using, the FACS and Action Units (AU). We verify our findings on the Extended Cohn-Kanade (CK+), NovaEmotions and FER2013 datasets. We apply these models to various tasks and tests using transfer learning, including cross-dataset validation and cross-task performance. Finally, we exploit the nature of the FER based CNN models for the detection of micro-expressions and achieve state-of-the-art accuracy using a simple long-short-term-memory (LSTM) recurrent neural network (RNN).
Forward citations
Cited by 2 Pith papers
-
Sparse Coding of Shape Trajectories for Facial Expression and Action Recognition
Applying intrinsic and extrinsic sparse coding and dictionary learning to Kendall shape trajectories yields vector-space time-series that perform competitively on 3D action and 2D facial expression recognition.
-
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
Milmer reports 96.72% four-class accuracy on DEAP by fusing facial frames selected via multiple instance learning with EEG tokens in a transformer, but the evaluation protocol is not fully described.
Discussion (0). Continue with ORCID to comment.