awesome-python-scientific-audio
by faroit
Curated list of python software and packages related to scientific research in audio
AI summary
Audio analysis library
A curated collection of Python packages and tools for scientific research in audio and music applications
- stars
- 1.6K
- forks
- 170
- watching
- 77
- awesome lists
- 6
- entries
- 100
What's in the list
100 links in 5 sections, with live GitHub stats.activeno commit in 2y
Audio Related Packages
audiolazy
Expressive Digital Signal Processing (DSP) package for Python
audioread
Cross-library (GStreamer + Core Audio + MAD + FFmpeg) audio decoding
- mutagen
Reads and writes all kind of audio metadata for various formats
- pyAV
PyAV is a Pythonic binding for FFmpeg or Libav
- (Py)Soundfile
Library based on libsndfile, CFFI, and NumPy
pySox
Wrapper for sox
stempeg
read/write of STEMS multistream audio
tinytag
reading music meta data of MP3, OGG, FLAC and Wave files
- acoustics
useful tools for acousticians
AudioTK
DSP filter toolbox (lots of filters)
- AudioTSM
real-time audio time-scale modification procedures
Gammatone
Gammatone filterbank implementation
- pyFFTW
Wrapper for FFTW(3)
- NSGT
Non-stationary gabor transform, constant-q
matchering
Automated reference audio mastering
MDCT
MDCT transform
- pydub
Manipulate audio with a simple and easy high level interface
- pytftb
Implementation of the MATLAB Time-Frequency Toolbox
pyroomacoustics
Room Acoustics Simulation (RIR generator)
PyRubberband
Wrapper for to do pitch-shifting and time-stretching
- PyWavelets
Discrete Wavelet Transform in Python
- Resampy
Sample rate conversion
- SFS-Python
Sound Field Synthesis Toolbox
- sound_field_analysis
Analyze, visualize and process sound field data recorded by spherical microphone arrays
- STFT
Standalone package for Short-Time Fourier Transform
- aubio
Feature extractor, written in C, Python interface
audioFlux
A library for audio and music analysis, feature extraction
audiolazy
Realtime Audio Processing lib, general purpose
- essentia
Music related low level and high level feature extractor, C++ based, includes Python bindings
python_speech_features
Common speech features for ASR
pyYAAFE
Python bindings for YAAFE feature extractor
speechpy
Library for Speech Processing and Recognition, mostly feature extraction for now
spafe
Python library for features extraction from audio files
audiomentations
Audio Data Augmentation
- muda
Musical Data Augmentation
pydiogment
Audio Data Augmentation
- aeneas
Forced aligner, based on MFCC+DTW, 35+ languages
deepspeech
Pretrained automatic speech recognition
gentle
Forced-aligner built on Kaldi
Parselmouth
Python interface to the phonetics and speech analysis, synthesis, and manipulation software
- persephone
Automatic phoneme transcription tool
pyannote.audio
Neural building blocks for speaker diarization
pyAudioAnalysis
² - Feature Extraction, Classification, Diarization
py-webrtcvad
Interface to the WebRTC Voice Activity Detector
pypesq
Wrapper for the PESQ score calculation
pystoi
Short Term Objective Intelligibility measure (STOI)
PyWorldVocoder
Wrapper for Morise's World Vocoder
- Montreal Forced Aligner
Forced aligner, based on Kaldi (HMM), English (others can be trained)
- SIDEKIT
Speaker and Language recognition
SpeechRecognition
Wrapper for several ASR engines and APIs, online and offline
- sed_eval
Evaluation toolbox for Sound Event Detection
cochlea
Inner ear models
- Brian2
Spiking neural networks simulator, includes cochlea model
Loudness
Perceived loudness, includes Zwicker, Moore/Glasberg model
- pyloudnorm
Audio loudness meter and normalization, implements ITU-R BS.1770-4
- Sound Field Synthesis Toolbox
Sound Field Synthesis Toolbox
commonfate
Common Fate Model and Transform
NTFLib
Sparse Beta-Divergence Tensor Factorization
- NUSSL
Holistic source separation framework including DSP methods and deep learning methods
- NIMFA
Several flavors of non-negative-matrix factorization
Catchy
Corpus Analysis Tools for Computational Hook Discovery
chord-detection
Algorithms for chord detection and key estimation
- Madmom
MIR packages with strong focus on beat detection, onset detection and chord recognition
- mir_eval
Common scores for various MIR tasks. Also includes bss_eval implementation
- msaf
Music Structure Analysis Framework
- librosa
General audio and music analysis
Kapre
Keras Audio Preprocessors
TorchAudio
PyTorch Audio Loaders
nnAudio
Accelerated audio processing using 1D convolution networks in PyTorch
- Music21
Toolkit for Computer-Aided Musicology
- Mido
Realtime MIDI wrapper
mingus
Advanced music theory and notation package with MIDI file and playback support
- Pretty-MIDI
Utility functions for handling MIDI data in a nice/intuitive way
Jupylet
Subtractive, additive, FM, and sample-based sound synthesis
- PYO
Realtime audio dsp engine
python-sounddevice
PortAudio wrapper providing realtime audio I/O with NumPy
ReTiSAR
Binarual rendering of streamed or IR-based high-order spherical microphone array signals
TimeSide (Beta)
high level audio analysis, imaging, transcoding, streaming and labelling
- beets
Music library manager and tagger
- musdb
Parse and process the MUSDB18 dataset
- medleydb
Parse audio + annotations
Soundcloud API
Wrapper for
- Youtube-Downloader
Download youtube videos (and the audio)
audiomate
Loading different types of audio datasets
- mirdata
Common loaders for Music Information Retrieval (MIR) datasets
- VamPy Host
Interface compiled vamp plugins
Tutorials
- Whirlwind Tour Of Python
fast-paced introduction to Python essentials, aimed at researchers and developers
- Introduction to Numpy and Scipy
Highly recommended tutorial, covers large parts of the scientific Python ecosystem
- Numpy for MATLAB® Users
Short overview of equivalent python functions for switchers
- MIR Notebooks
collection of instructional iPython Notebooks for music information retrieval (MIR)
Selected Topics in Audio Signal Processing
Exercises as iPython notebooks
- Live-coding a music synthesizer
Live-coding video showing how to use the SoundDevice library to reproduce realistic sounds.
Books
Python Data Science Handbook
Jake Vanderplas, Excellent Book and accompanying tutorial notebooks
- Fundamentals of Music Processing
Meinard Müller, comes with Python exercises
Scientific Papers
- Python for audio signal processing
John C. Glover, Victor Lazzarini and Joseph Timoney, Linux Audio Conference 2011
- librosa: Audio and Music Signal Analysis in Python
, - Brian McFee, Colin Raffel, Dawen Liang, Daniel P.W. Ellis, Matt McVicar, Eric Battenberg, Oriol Nieto, Scipy 2015
- pyannote.audio: neural building blocks for speaker diarization
, - Hervé Bredin, Ruiqing Yin, Juan Manuel Coria, Gregory Gelly, Pavel Korshunov, Marvin Lavechin, Diego Fustes, Hadrien Titeux, Wassim Bouaziz, Marie-Philippe Gill, ICASSP 2020
Other Resources
- Coursera Course
Audio Signal Processing, Python based course from UPF of Barcelona and Stanford University
- Digital Signal Processing Course
Masters Course Material (University of Rostock) with many Python examples
- Slack Channel
Music Information Retrieval Community
Nothing in this list matches your filter.
Featured in 6 awesome lists
Each link jumps to the spot where the list mentions awesome-python-scientific-audio.
More related projects
anishathalye/neural-style5.5K
softcatala/whisper-ctranslate2938
scipy/scipy13.2K
m-bain/whisperx12.9K
avsystem/anjay191
danshapero/icepack-py1
yoggy/sendosc154
farama-foundation/arcade-learning-environment2.2K
ifm/ifm3d112
yuki-koyama/mathtoolbox266
ml-gde/e2e-tflite-tutorials133
alexandre01/ultimatelabeling316
unslothai/hyperlearn1.9K
ibm/max-speech-to-text-converter76