Lipreading-DenseNet3D

Lip movement analyzer

A software implementation of a deep learning model designed to understand lip movements in videos

DenseNet3D Model In "LRW-1000: A Naturally-Distributed Large-Scale Benchmark for Lip Reading in the Wild", https://arxiv.org/abs/1810.06990

GitHub

117 stars
6 watching
21 forks
Language: Python
last commit: almost 6 years ago
arxivdeeplearninglipreadingpytorch

Related projects:

RepositoryDescriptionStars
astorfi/lip-reading-deeplearningDeep learning-based system for recognizing speech from lip movements using 3D convolutional neural networks.1,840
millionintegrals/velA collection of modular deep learning components that can be easily configured and reused in various applications.276
hualin95/deeplab-v3plusA high-performance PyTorch implementation of semantic image segmentation using a custom encoder-decoder architecture.334
astorfi/3d-convolutional-speaker-recognitionDevelops deep learning models using 3D convolutional neural networks for speaker verification tasks783
devendrachaplot/deeprl-groundingTrains an RL agent to execute natural language instructions in a 3D environment using a combination of A3C and gated attention mechanisms.237
vita-epfl/crowdnavDevelops robot navigation policies in crowded spaces using reinforcement learning and attention mechanisms.607
codeslake/pvdnetAn open-source implementation of a deep learning model for video deblurring and motion estimation.114
dvlab-research/prompt-highlighterAn interactive control system for text generation in multi-modal language models135
clementpinard/sfmlearner-pytorchPytorch implementation of unsupervised depth and ego-motion learning from video sequences1,022
foamliu/deep-image-matting-pytorchAn implementation of deep image matting in PyTorch using a neural network architecture.821
engineering-course/lip_sslA deep learning framework for human parsing that learns to detect human structures without explicit joint labeling.229
dvlab-research/lisaA system that uses large language models to generate segmentation masks for images based on complex queries and world knowledge.1,923
deepwisdom/autodlAutomated deep learning algorithm that performs feature engineering, model selection, and hyperparameter tuning without human intervention.1,140
vita-epfl/monolocoA software framework for 3D vision and computer vision tasks using deep learning and 2D keypoints.431
nvidia-merlin/nvtabularA library that provides a high-level abstraction for feature engineering and preprocessing of tabular data to accelerate deep learning recommender systems on GPUs.1,057