跳到主要导航 跳到搜索 跳到主要内容

Deep Video-Based Performance Synthesis from Sparse Multi-View Capture

  • Mingjia Chen
  • , Changbo Wang*
  • , Ligang Liu
  • *此作品的通讯作者
  • East China Normal University
  • University of Science and Technology of China

科研成果: 期刊稿件文章同行评审

摘要

We present a deep learning based technique that enables novel-view videos of human performances to be synthesized from sparse multi-view captures. While performance capturing from a sparse set of videos has received significant attention, there has been relatively less progress which is about non-rigid objects (e.g., human bodies). The rich articulation modes of human body make it rather challenging to synthesize and interpolate the model well. To address this problem, we propose a novel deep learning based framework that directly predicts novel-view videos of human performances without explicit 3D reconstruction. Our method is a composition of two steps: novel-view prediction and detail enhancement. We first learn a novel deep generative query network for view prediction. We synthesize novel-view performances from a sparse set of just five or less camera videos. Then, we use a new generative adversarial network to enhance fine-scale details of the first step results. This opens up the possibility of high-quality low-cost video-based performance synthesis, which is gaining popularity for VA and AR applications. We demonstrate a variety of promising results, where our method is able to synthesis more robust and accurate performances than existing state-of-the-art approaches when only sparse views are available.

源语言英语
页(从-至)543-554
页数12
期刊Computer Graphics Forum
38
7
DOI
出版状态已出版 - 1 10月 2019

指纹

探究 'Deep Video-Based Performance Synthesis from Sparse Multi-View Capture' 的科研主题。它们共同构成独一无二的指纹。

引用此