G3AN: Disentangling Appearance and Motion for Video Generation

Yaohui Wang; Piotr Bilinski; Francois Bremond; Antitza Dantcheva

arXiv:1912.05523·cs.CV·June 16, 2020·1 cites

G3AN: Disentangling Appearance and Motion for Video Generation

Yaohui Wang, Piotr Bilinski, Francois Bremond, Antitza Dantcheva

PDF

Open Access 1 Repo 1 Video

TL;DR

G3AN is a novel spatio-temporal generative model that disentangles appearance and motion to generate realistic human videos, outperforming existing methods on multiple datasets.

Contribution

The paper introduces G3AN, a three-stream generator that effectively models and disentangles appearance and motion in video generation.

Findings

01

Outperforms state-of-the-art on facial expression datasets

02

Achieves significant improvements on human action datasets

03

Successfully learns disentangled appearance and motion representations

Abstract

Creating realistic human videos entails the challenge of being able to simultaneously generate both appearance, as well as motion. To tackle this challenge, we introduce G $^{3}$ AN, a novel spatio-temporal generative model, which seeks to capture the distribution of high dimensional video data and to model appearance and motion in disentangled manner. The latter is achieved by decomposing appearance and motion in a three-stream Generator, where the main stream aims to model spatio-temporal consistency, whereas the two auxiliary streams augment the main stream with multi-scale appearance and motion features, respectively. An extensive quantitative and qualitative analysis shows that our model systematically and significantly outperforms state-of-the-art methods on the facial expression datasets MUG and UvA-NEMO, as well as the Weizmann and UCF101 datasets on human action. Additional…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

wyhsirius/g3an-project
pytorchOfficial

Videos

G3AN: Disentangling Appearance and Motion for Video Generation· youtube

Taxonomy

TopicsGenerative Adversarial Networks and Image Synthesis · Face recognition and analysis · Cinema and Media Studies