Awesome_Audio-driven_Talking-Face-Generation

Talking face generator resource list

A curated list of resources for generating talking faces based on audio input

A curated list of resources of audio-driven talking face generation

GitHub

133 stars
2 watching
11 forks
last commit: about 4 years ago
Linked from 1 awesome list

audio-driven-talking-facecontrollable-generationgenerative-adversarial-networkspaperlistspeech-driven-talking-facetalking-face-generation

Awesome Audio-driven Talking Face Generation / 2D Encoder-Decoder Based

demoStyleHEAT: One-Shot High-Resolution Editable Talking Face Generation via Pre-trained StyleGAN [F Yin 2022] [arXiv]
demoPose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation [Hang Zhou 2021] [CVPR]
project page167over 3 years agoSpeech Driven Talking Face Generation from a Single Image and an Emotion Condition [SE Eskimez 2021] [arXiv]
demoHeadGAN: Video-and-Audio-Driven Talking Head Synthesis [MC Doukas 2021] [arXiv]
demoA Lip Sync Expert Is All You Need for Speech to Lip Generation In The Wild [K R Prajwal 2020] [ACMMM]
demoLearning Individual Speaking Styles for Accurate Lip to Speech Synthesis [KR Prajwal 2020] [CVPR]
demoRobust One Shot Audio to Video Generation [N Kumar 2020] [CVPRW]
demoTalking Face Generation by Adversarially Disentangled Audio-Visual Representation [Hang Zhou 2019] [AAAI]
demoTalking face generation by conditional recurrent adversarial network [Yang Song 2019] [IJCAI]
demoRealistic Speech-Driven Facial Animation with GANs [Konstantinos Vougioukas 2019] [IJCV]
demoLip Movements Generation at a Glance [Lele Chen 2018] [ECCV]
demoX2Face: A network for controlling face generation using images, audio, and pose codes [Olivia Wiles 2018] [ECCV]
demoGenerative Adversarial Talking Head: Bringing Portraits to Life with a Weakly Supervised Neural Network [HX Pham 2018] [arXiv]
demoYou said that? [Chung 2017] [BMVC]

Awesome Audio-driven Talking Face Generation / Landmark Based

demoLive Speech Portraits: Real-Time Photorealistic Talking-Head Animation [YUANXUN LU 2021] [SIGGRAPH]
demo278almost 5 years agoImitating Arbitrary Talking Style for Realistic Audio-Driven Talking Face Synthesis [H Wu 2021] [ACMMM]
demoMakeItTalk: Speaker-Aware Talking-Head Animation [YANG ZHOU 2020] [SIGGRAPH]
demoHierarchical Cross-Modal Talking Face Generation with Dynamic Pixel-Wise Loss [Lele Chen 2019] [CVPR]
demoSynthesizing Obama: learning lip sync from audio [SUPASORN SUWAJANAKORN 2017] [SIGGRAPH]

Awesome Audio-driven Talking Face Generation / 3D Model Based

demoEverybody’s Talkin’: Let Me Talk as You Want [Linsen Song 2022] [TIFS]
demoOne-shot Talking Face Generation from Single-speaker Audio-Visual Correlation Learning [Suzhen Wang 2022] [AAAI]
demoFaceFormer: Speech-Driven 3D Facial Animation with Transformers [Y Fan 2022] [CVPR]
demoIterative Text-based Editing of Talking-heads Using Neural Retargeting [Xinwei Yao 2021] [ICML]
demoAD-NeRF: Audio Driven Neural Radiance Fields for Talking Head Synthesis [Yudong Guo 2021] [ICCV]
demoAudio-driven emotional video portraits [X Ji 2021] [CVPR]
demoFACIAL: Synthesizing Dynamic Talking Face with Implicit Attribute Learning [C Zhang 2021] [ICCV]
demoFlow-guided One-shot Talking Face Generation with a High-resolution Audio-visual Dataset [Z Zhang 2021] [CVPR]
demoAudio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion [Suzhen Wang 2021] [IJCAI]
demoMeshTalk: 3D Face Animation from Speech using Cross-Modality Disentanglement [A Richard 2021] [ICCV]
demoWrite-a-speaker: Text-based Emotional and Rhythmic Talking-head Generation [L Li 2021] [AAAI]
demoText2Video: Text-driven Talking-head Video Synthesis with Phonetic Dictionary [S Zhang 2021 ] [ICASSP]
demoNeural Voice Puppetry: Audio-driven Facial Reenactment [Justus Thies 2020] [ECCV]
projectpage740almost 3 years agoAudio-driven Talking Face Video Generation with Learning-based Personalized Head Pose [Ran Yi 2020] [arXiv]
demoTalking-head Generation with Rhythmic Head Motion [Lele Chen 2020] [ECCV]
demoAudio- and Gaze-driven Facial Animation of Codec Avatars [A Richard 2020] [arXiv]
demoText-based editing of talking-head video [OHAD FRIED 2019] [arXiv]
demoCapture, Learning, and Synthesis of 3D Speaking Styles [D Cudeiro 2019] [CVPR]
demoVisemenet: audio-driven animator-centric speech animation [YANG ZHOU 2018] [TOG]
demoAudio-Driven Facial Animation by Joint End-to-End Learning of Pose and Emotion [TERO KARRAS 2017] [TOG]
demoA deep learning approach for generalized speech animation [SARAH TAYLOR 2017] [SIGGRAPH]
demoJALI: An Animator-Centric Viseme Model for Expressive Lip Synchronization [Pif Edwards 2016] [SIGGRAPH]

Awesome Audio-driven Talking Face Generation / Datasets

project pageGRID 2006
project pageTCD-TIMIT 2015
project pageLRW 2016
project pageMODALITY 2017
project pageVoxceleb1 2017
project pageVoxceleb2 2018
project pageLRS2-BBC 2018
project pageLRS3-TED 2018
project page349over 2 years agoHDTF 2020
project page379almost 4 years agoCREMA-D 2014
project pageMSP-IMPROV 2016
project pageRAVDESS 2018
project pageMELD 2018
project pageMEAD 2020
project page117almost 6 years agoLRW-1000 2018

Backlinks from these awesome lists: