arXiv · 1803.07461
Speech-Driven Facial Reenactment Using Conditional Generative Adversarial Networks
Abstract
We present a novel approach to generating photo-realistic images of a face with accurate lip sync, given an audio input. By using a recurrent neural network, we achieved mouth landmarks based on audio features. We exploited the power of conditional generative adversarial networks to produce highly-realistic face conditioned on a set of landmarks. These two networks together are capable of producing a sequence of natural faces in sync with an input audio track.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Seyed Ali Jalalifar, Hosein Hasani, Hamid Aghajan. 2018-03-20. Speech-Driven Facial Reenactment Using Conditional Generative Adversarial Networks. https://arxiv.org/abs/1803.07461
Cite the original work for its findings. Save a collection to share your selection of sources.