python_speech_features

Speech feature extractor

A library that computes speech features commonly used in Automatic Speech Recognition (ASR) systems.

This library provides common speech features for ASR including MFCCs and filterbank energies.

GitHub

2k stars
87 watching
617 forks
Language: Python
last commit: almost 5 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
jameslyons/matlab_speech_featuresA set of speech feature extraction functions for various machine learning applications.43
astorfi/speechpyProvides tools and libraries for extracting speech features from audio data.881
tyiannak/pyaudioanalysisA comprehensive Python library for feature extraction, classification, segmentation, and applications of audio data.5,918
pytorch/audioA PyTorch module providing tools and functions for audio signal processing2,561
superkogito/spafeA Python library for extracting audio features from mono audio files using various filter banks and spectrogram algorithms.461
rf5/transfusion-asrAn ASR project that uses diffusion models to transcribe speech76
peak1995/speech-enhancement-dspThis repository provides MATLAB implementations of traditional speech enhancement techniques including spectral subtraction, Wiener filtering, and Kalman filtering.84
awni/speechA PyTorch implementation of end-to-end speech recognition models.756
bmcfee/pyrubberbandProvides a lightweight Python wrapper for audio processing tasks167
belangeo/pyoA Python module for digital signal processing and audio synthesis, allowing users to create complex audio chains in real-time.1,329
iver56/audiomentationsLibrary for audio data augmentation used in machine learning1,903
cpjku/madmomA Python audio signal processing library used in music information retrieval tasks.1,366
linto-ai/whisper-timestampedAn extension to the Whisper speech recognition model that adds word-level timestamps and confidence scores.2,121
vocalpy/vakA Python framework for training and applying neural networks to acoustic communication research78
r9y9/tacotron_pytorchAn implementation of Tacotron speech synthesis model using PyTorch.309