G3AN: Disentangling Appearance and Motion for Video Generation

被引:52
作者
Wang, Yaohui [1 ,2 ]
Bilinski, Piotr [3 ]
Bremond, Francois [1 ,2 ]
Dantcheva, Antitza [1 ,2 ]
机构
[1] INRIA, Sophia Antipolis, France
[2] Univ Cote dAzur, Nice, France
[3] Univ Warsaw, Warsaw, Poland
来源
2020 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) | 2020年
关键词
D O I
10.1109/CVPR42600.2020.00531
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Creating realistic human videos entails the challenge of being able to simultaneously generate both appearance, as well as motion. To tackle this challenge, we introduce G(3)AN, a novel spatio-temporal generative model, which seeks to capture the distribution of high dimensional video data and to model appearance and motion in disentangled manner. The latter is achieved by decomposing appearance and motion in a three-stream Generator, where the main stream aims to model spatio-temporal consistency, whereas the two auxiliary streams augment the main stream with multi-scale appearance and motion features, respectively. An extensive quantitative and qualitative analysis shows that our model systematically and significantly outperforms state-of-the-art methods on the facial expression datasets MUG and UvA-NEMO, as well as the Weizmann and UCF101 datasets on human action. Additional analysis on the learned latent representations confirms the successful decomposition of appearance and motion. Source code and pre-trained models are publicly available(1).
引用
收藏
页码:5263 / 5272
页数:10
相关论文
共 45 条
[1]  
Aifanti N., 2010, 11 INT WORKSH IM AN, P1, DOI DOI 10.1371/JOURNAL.PONE.0009715
[2]   Guided Image-to-Image Translation with Bi-Directional Feature Transformation [J].
AlBahar, Badour ;
Huang, Jia-Bin .
2019 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2019), 2019, :9015-9024
[3]  
[Anonymous], 2018, ECCV
[4]  
[Anonymous], 2016, NIPS
[5]  
[Anonymous], 2018, ICML
[6]  
[Anonymous], INT C LEARNING REPRE
[7]  
[Anonymous], 2018, ECCV
[8]  
[Anonymous], 2016, P ADV NEURAL INFORM
[9]  
[Anonymous], 2017, ICCV
[10]  
[Anonymous], 2012, ECCV