Audio-driven-TalkingFace-HeadPose

Talking Face Generator

Generates talking face videos based on audio signals and personalized head poses.

Code for "Audio-driven Talking Face Video Generation with Learning-based Personalized Head Pose" (Arxiv 2020) and "Predicting Personalized Head Movement From Short Video and Speech Signal" (TMM 2022)

GitHub

740 stars
25 watching
146 forks
Language: Python
last commit: almost 3 years ago

Related projects:

RepositoryDescriptionStars
mrzzm/hdtfA project providing a dataset and code for generating talking faces with high-resolution audio-visual data349
eeskimez/emotalkingfaceA system that generates talking faces from images and speech with different emotions.167
pkhungurn/talking-head-anime-demoCreates anime characters with realistic head movements from single images or webcam feeds using deep learning and computer vision techniques.2,001
akanimax/t2fGenerates images of human faces based on textual descriptions using deep learning techniques.548
zhanglonghao1992/one-shot_free-view_neural_talking_head_synthesisAn implementation of neural talking head synthesis for video conferencing, allowing for one-shot creation of realistic face movements.807
dmitryulyanov/ageThis repository provides code for training Generative Adversarial Networks (GANs) for various image datasets, including face generation.285
wuhaozhe/style_avatarGenerates stylized talking faces and videos using deep learning models278
pavitrakumar78/anime-face-gan-kerasA GAN-based system to generate anime faces using a custom dataset198
chrisdonahue/waveganAn open-source machine learning algorithm for generating raw audio waveforms from raw data1,334
a312863063/seeprettyface-ganerator-dongmanA Python implementation of a StyleGAN-based anime face generator151
lelechen63/atvgnetThis repository provides implementations of neural networks used in cross-modal talking face generation258
yi-ming-qian/roofganA tool for generating realistic roof models using deep learning techniques42
yitong91/storyganA framework for generating images that describe stories using deep learning techniques233
iigroup/mm-celeba-hq-datasetA large-scale dataset for training and evaluating algorithms for text-driven face generation and understanding tasks.223
soroushmehr/samplernn_iclr2017An unconditional end-to-end neural audio generation model utilizing a recurrent neural network architecture.537