Pith. sign in

REVIEW 1 cited by

Capturing Long-term Temporal Dependencies with Convolutional Networks for Continuous Emotion Recognition

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1708.07050 v1 pith:73OYSY5H submitted 2017-08-23 cs.SD cs.AI

classification cs.SDcs.AI
keywords emotionarchitecturescontinuousoutputrecognitiondependencieslong-termperformance
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

The goal of continuous emotion recognition is to assign an emotion value to every frame in a sequence of acoustic features. We show that incorporating long-term temporal dependencies is critical for continuous emotion recognition tasks. To this end, we first investigate architectures that use dilated convolutions. We show that even though such architectures outperform previously reported systems, the output signals produced from such architectures undergo erratic changes between consecutive time steps. This is inconsistent with the slow moving ground-truth emotion labels that are obtained from human annotators. To deal with this problem, we model a downsampled version of the input signal and then generate the output signal through upsampling. Not only does the resulting downsampling/upsampling network achieve good performance, it also generates smooth output trajectories. Our method yields the best known audio-only performance on the RECOLA dataset.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Controlling for Confounders in Multimodal Emotion Classification via Adversarial Learning

    cs.LG 2019-08 conditional novelty 6.0 of 10

    Adversarial training that removes stress-related signals from emotion representations improves cross-dataset emotion recognition in several, but not all, test conditions.

Pith tools