AudioLDM

Audio Generator

A Python-based audio generation tool that can produce speech, sound effects, music, and more, using text as input or guided by user description.

AudioLDM: Generate speech, sound effects, music and beyond, with text.

GitHub

2k stars
42 watching
225 forks
Language: Python
last commit: almost 2 years ago
audio-generation

Related projects:

RepositoryDescriptionStars
grame-cncm/faustA functional programming language for real-time signal processing and synthesis2,605
suno-ai/barkA text-to-audio model that generates realistic speech and other audio36,433
aigc-audio/audiogptAn audio processing toolkit that provides pre-trained models and tools for tasks like speech synthesis, music generation, sound detection, and talking head creation.10,061
lucidrains/musiclm-pytorchImplementation of Google's MusicLM model for music generation using attention networks and text-conditioning.3,189
facebookresearch/audiocraftA deep learning library for generating high-quality audio21,134
ibm/max-audio-sample-generatorA tool to generate audio samples based on input commands and lo-fi instrumental music tracks.22
jiaaro/pydubA Python library for manipulating and editing audio files9,024
superkogito/pydiogmentA Python library for generating multiple audio files based on a starting mono audio file with various effects such as speed change, tone alteration and noise addition.83
jasonppy/voicecraftA neural codec model for speech editing and text-to-speech synthesis in real-time, using few seconds of reference audio.7,744
juandagilc/audio-effectsA collection of audio effects plugins implemented from a book and contributing examples757
oobabooga/text-generation-webuiA web-based interface for generating text using large language models41,123
mubertai/mubert-text-to-musicGenerates music based on user input prompts using the Mubert API2,738
gl326/bard-audioAn audio engine for Game Maker Studio 2 designed to facilitate good audio implementation.38
rustaudio/cpalA cross-platform audio I/O library in pure Rust2,772
nvidia/waveglowGenerates high-quality speech from mel-spectrograms using a flow-based network architecture2,294