audio

Audio toolkit

A PyTorch module providing tools and functions for audio signal processing

Data manipulation and transformation for audio signal processing, powered by PyTorch

GitHub

3k stars
73 watching
659 forks
Language: Python
last commit: almost 2 years ago
Linked from 4 awesome lists

audioaudio-processingiomachine-learningpythonpytorchspeech

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
facebookresearch/audiocraftA deep learning library for generating high-quality audio21,134
kinwaicheuk/nnaudioAn audio processing toolkit using PyTorch convolutional neural networks to generate spectrograms from raw audio data1,036
pytorch/pytorchA Python library providing tensors and dynamic neural networks with strong GPU acceleration84,978
lucidrains/musiclm-pytorchImplementation of Google's MusicLM model for music generation using attention networks and text-conditioning.3,189
archinetai/audio-diffusion-pytorchAn audio generation library that uses diffusion models to produce high-quality audio samples from noise or text input1,975
tyiannak/pyaudioanalysisA comprehensive Python library for feature extraction, classification, segmentation, and applications of audio data.5,918
pyannote/pyannote-audioA toolkit for speaker diarization using PyTorch and speech activity detection.6,508
pytorch/torchtuneA PyTorch library for easily authoring and experimenting with large language models4,479
nvidia/tacotron2This PyTorch implementation provides a toolkit for speech synthesis using a deep neural network architecture.5,123
deepsound-project/samplernn-pytorchAn implementation of an audio generation model using PyTorch290
nvidia/apexTools for streamlined mixed precision and distributed training in PyTorch8,460
libaudioflux/audiofluxA deep learning tool library for extracting features from audio signals.2,940
pytorch/torchtitanA native PyTorch library for training large language models using distributed parallelism and optimization techniques.2,765
microsoft/torchgeoProvides tools and pre-trained models for working with geospatial data in machine learning applications3,083
lcav/pyroomacousticsSoftware package for rapid development and testing of audio array processing algorithms in indoor applications1,480