aeneas

Alignment tool

Automatically synchronizes text and audio to create a synchronized alignment

aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)

GitHub

3k stars
72 watching
236 forks
Language: Python
last commit: over 2 years ago
Linked from 1 awesome list

alignmentaudioclidtwespeakespeak-ngfestivalffmpegforced-alignmentlinuxmacosnlppythonsmilspeechsrttexttext-to-speechttswindows

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
aubio/aubioA comprehensive library for audio and music analysis and processing.3,336
pettarin/forced-alignment-toolsA collection of tools and resources for computing forced alignments between audio files and transcripts.878
pytorch/audioA PyTorch module providing tools and functions for audio signal processing2,561
montrealcorpustools/montreal-forced-alignerA command-line utility for aligning audio data with written text based on pronunciation rules.1,364
machinalis/yalignAutomates the process of extracting parallel sentences from comparable corpora to aid in statistical machine translation127
tp7/sushiAutomates subtitle syncing by comparing audio patterns to align subtitles with different video sources.649
eleutherai/pythiaAnalyzing knowledge development and evolution in large language models during training2,309
laion-ai/clapA library for learning audio embeddings from text and audio data using contrastive language-audio pretraining1,457
smacke/ffsubsyncAutomatically synchronizes subtitles with video6,879
aixander/realtime_pyaudio_fftAn audio analysis tool that extracts and visualizes features from live audio streams using FFTs.976
iver56/audiomentationsLibrary for audio data augmentation used in machine learning1,903
eli64s/readme-aiAutomates the generation of comprehensive README files using AI-powered language models.1,665
prosodylab/prosodylab-alignerTools for aligning laboratory speech production data to forced audio alignment using HTK and SoX.333
lowerquality/gentleA tool for aligning speech with text by analyzing audio and providing an output transcript1,471
pyannote/pyannote-audioA toolkit for speaker diarization using PyTorch and speech activity detection.6,508