Facial Expression Translation using Landmark Guided GANs

Hao Tang; Nicu Sebe

arXiv:2209.02136·cs.CV·September 7, 2022

Facial Expression Translation using Landmark Guided GANs

Hao Tang, Nicu Sebe

PDF

Open Access 1 Repo

TL;DR

This paper introduces LandmarkGAN, a novel GAN-based method that uses facial landmarks to translate facial expressions from a single image, outperforming existing keypoint-guided approaches.

Contribution

LandmarkGAN is the first to explicitly incorporate landmark information for expression translation using only a single image, with a two-stage end-to-end training process.

Findings

01

Outperforms state-of-the-art methods on four datasets.

02

Requires only a single image for expression translation.

03

Effective landmark-guided translation in diverse poses and backgrounds.

Abstract

We propose a simple yet powerful Landmark guided Generative Adversarial Network (LandmarkGAN) for the facial expression-to-expression translation using a single image, which is an important and challenging task in computer vision since the expression-to-expression translation is a non-linear and non-aligned problem. Moreover, it requires a high-level semantic understanding between the input and output images since the objects in images can have arbitrary poses, sizes, locations, backgrounds, and self-occlusions. To tackle this problem, we propose utilizing facial landmark information explicitly. Since it is a challenging problem, we split it into two sub-tasks, (i) category-guided landmark generation, and (ii) landmark-guided expression-to-expression translation. Two sub-tasks are trained in an end-to-end fashion that aims to enjoy the mutually improved benefits from the generated…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ha0tang/landmarkgan
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsFace recognition and analysis · Generative Adversarial Networks and Image Synthesis · Speech and Audio Processing