Improvement in Facial Emotion Recognition using Synthetic Data Generated   by Diffusion Model

Arnab Kumar Roy; Hemant Kumar Kathania; Adhitiya Sharma

arXiv:2411.10863·cs.CV·November 19, 2024

Improvement in Facial Emotion Recognition using Synthetic Data Generated by Diffusion Model

Arnab Kumar Roy, Hemant Kumar Kathania, Adhitiya Sharma

PDF

Open Access 1 Repo

TL;DR

This paper demonstrates that using synthetic facial emotion data generated by diffusion models significantly improves the accuracy of emotion recognition systems, especially in imbalanced datasets.

Contribution

It introduces a novel data augmentation approach using diffusion models to enhance FER performance with the ResEmoteNet model.

Findings

01

Achieved 96.47% accuracy on FER2013 dataset

02

Achieved 99.23% accuracy on RAF-DB dataset

03

Significant performance improvements over baseline models

Abstract

Facial Emotion Recognition (FER) plays a crucial role in computer vision, with significant applications in human-computer interaction, affective computing, and areas such as mental health monitoring and personalized learning environments. However, a major challenge in FER task is the class imbalance commonly found in available datasets, which can hinder both model performance and generalization. In this paper, we tackle the issue of data imbalance by incorporating synthetic data augmentation and leveraging the ResEmoteNet model to enhance the overall performance on facial emotion recognition task. We employed Stable Diffusion 2 and Stable Diffusion 3 Medium models to generate synthetic facial emotion data, augmenting the training sets of the FER2013 and RAF-DB benchmark datasets. Training ResEmoteNet with these augmented datasets resulted in substantial performance improvements,…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ArnabKumarRoy02/ResEmoteNet
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsConsumer Perception and Purchasing Behavior · Face and Expression Recognition

MethodsDiffusion