An audiovisual and contextual approach for categorical and continuous   emotion recognition in-the-wild

Panagiotis Antoniadis; Ioannis Pikoulis; Panagiotis P. Filntisis,; Petros Maragos

arXiv:2107.03465·cs.CV·November 4, 2022

An audiovisual and contextual approach for categorical and continuous emotion recognition in-the-wild

Panagiotis Antoniadis, Ioannis Pikoulis, Panagiotis P. Filntisis,, Petros Maragos

PDF

1 Repo

TL;DR

This paper presents a multi-modal, multi-stream deep learning framework for in-the-wild emotion recognition that incorporates facial, bodily, and contextual features to improve robustness under challenging conditions.

Contribution

It introduces a novel multi-stream CNN-RNN model leveraging audiovisual and contextual cues, demonstrating improved emotion recognition performance in unconstrained environments.

Findings

01

Multi-modal approach enhances recognition accuracy.

02

Inclusion of body and scene context improves robustness.

03

Model outperforms existing methods on Aff-Wild2 dataset.

Abstract

In this work we tackle the task of video-based audio-visual emotion recognition, within the premises of the 2nd Workshop and Competition on Affective Behavior Analysis in-the-wild (ABAW2). Poor illumination conditions, head/body orientation and low image resolution constitute factors that can potentially hinder performance in case of methodologies that solely rely on the extraction and analysis of facial features. In order to alleviate this problem, we leverage both bodily and contextual features, as part of a broader emotion recognition framework. We choose to use a standard CNN-RNN cascade as the backbone of our proposed model for sequence-to-sequence (seq2seq) learning. Apart from learning through the RGB input modality, we construct an aural stream which operates on sequences of extracted mel-spectrograms. Our extensive experiments on the challenging and newly assembled Aff-Wild2…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

PanosAntoniadis/NTUA-ABAW2021
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.