encodec

Audio codec

A deep learning-based audio codec that supports high-fidelity neural audio compression.

State-of-the-art deep learning based audio codec supporting both mono 24 kHz audio and stereo 48 kHz audio.

GitHub

4k stars
57 watching
309 forks
Language: Python
last commit: over 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
facebookresearch/audiocraftA deep learning library for generating high-quality audio21,134
xiph/rnnoiseA deep learning-based audio noise reduction system using recurrent neural networks4,191
intel/neural-compressorTools and techniques for optimizing large language models on various frameworks and hardware platforms.2,257
facebookresearch/audio2photorealGenerating photorealistic avatars from audio2,715
libaudioflux/audiofluxA deep learning tool library for extracting features from audio signals.2,940
facebookresearch/demucsA deep learning model that separates multiple audio sources from mixed music tracks8,453
enhuiz/vall-eAn implementation of VALL-E in PyTorch for text-to-speech synthesis2,970
deep-floyd/ifA text-to-image synthesis model with a modular design, utilizing a frozen text encoder and cascaded pixel diffusion modules to generate photorealistic images.7,699
lucidrains/musiclm-pytorchImplementation of Google's MusicLM model for music generation using attention networks and text-conditioning.3,189
mubertai/mubert-text-to-musicGenerates music based on user input prompts using the Mubert API2,738
nvidia/waveglowGenerates high-quality speech from mel-spectrograms using a flow-based network architecture2,294
oxford-cs-deepnlp-2017/lecturesAn open-source repository containing lecture slides and course materials for an advanced natural language processing course.15,702
soerenab/audiomnistThis project provides an implementation of a deep learning framework to classify audio signals and offers insights into the model's decision-making process using Explainable Artificial Intelligence (AI) techniques.351
csteinmetz1/micro-tcnA software framework for efficient modeling of analog audio dynamic range compression using neural networks150
pyannote/pyannote-audioA toolkit for speaker diarization using PyTorch and speech activity detection.6,508