Deep Video-Based Performance Cloning

K. Aberman, M. Shi, J. Liao, D. Lischinski, B. Chen, D. Cohen-Or

Research output: Contribution to journalArticlepeer-review

61 Scopus citations


We present a new video-based performance cloning technique. After training a deep generative network using a reference video capturing the appearance and dynamics of a target actor, we are able to generate videos where this actor reenacts other performances. All of the training data and the driving performances are provided as ordinary video segments, without motion capture or depth information. Our generative model is realized as a deep neural network with two branches, both of which train the same space-time conditional generator, using shared weights. One branch, responsible for learning to generate the appearance of the target actor in various poses, uses paired training data, self-generated from the reference video. The second branch uses unpaired data to improve generation of temporally coherent video renditions of unseen pose sequences. Through data augmentation, our network is able to synthesize images of the target actor in poses never captured by the reference video. We demonstrate a variety of promising results, where our method is able to generate temporally coherent videos, for challenging scenarios where the reference and driving videos consist of very different dance performances.

Original languageAmerican English
Pages (from-to)219-233
Number of pages15
JournalComputer Graphics Forum
Issue number2
StatePublished - May 2019

Bibliographical note

Publisher Copyright:
© 2019 The Author(s) Computer Graphics Forum © 2019 The Eurographics Association and John Wiley & Sons Ltd. Published by John Wiley & Sons Ltd.


  • CCS Concepts
  • Neural networks
  • • Computing methodologies → Image-based rendering


Dive into the research topics of 'Deep Video-Based Performance Cloning'. Together they form a unique fingerprint.

Cite this