Awesome Lists

awesome-python-scientific-audio

by faroit

awesome listpushed about 3 years ago

Curated list of python software and packages related to scientific research in audio

AI summary

Audio analysis library

A curated collection of Python packages and tools for scientific research in audio and music applications

stars
1.6K
forks
170
watching
77
awesome lists
6
entries
100
View on GitHub

Embed the badge

Show how many awesome lists link to your project. The count updates automatically.

Awesome Lists badge
Markdown
[![Awesome Lists Badge](https://awesome.facts.dev/shield/faroit/awesome-python-scientific-audio/links.svg)](https://awesome.facts.dev/awesome/faroit/awesome-python-scientific-audio)
HTML
<a href="https://awesome.facts.dev/awesome/faroit/awesome-python-scientific-audio"><img src="https://awesome.facts.dev/shield/faroit/awesome-python-scientific-audio/links.svg" alt="Awesome Lists Badge" /></a>
Image URL
https://awesome.facts.dev/shield/faroit/awesome-python-scientific-audio/links.svg

What's in the list

100 links in 5 sections, with live GitHub stats.activeno commit in 2y

  • audiolazy

    Expressive Digital Signal Processing (DSP) package for Python

  • audioread

    Cross-library (GStreamer + Core Audio + MAD + FFmpeg) audio decoding

  • mutagen

    Reads and writes all kind of audio metadata for various formats

  • pyAV

    PyAV is a Pythonic binding for FFmpeg or Libav

  • (Py)Soundfile

    Library based on libsndfile, CFFI, and NumPy

  • pySox

    Wrapper for sox

  • stempeg

    read/write of STEMS multistream audio

  • tinytag

    reading music meta data of MP3, OGG, FLAC and Wave files

  • acoustics

    useful tools for acousticians

  • AudioTK

    DSP filter toolbox (lots of filters)

  • AudioTSM

    real-time audio time-scale modification procedures

  • Gammatone

    Gammatone filterbank implementation

  • pyFFTW

    Wrapper for FFTW(3)

  • NSGT

    Non-stationary gabor transform, constant-q

  • matchering

    Automated reference audio mastering

  • MDCT

    MDCT transform

  • pydub

    Manipulate audio with a simple and easy high level interface

  • pytftb

    Implementation of the MATLAB Time-Frequency Toolbox

  • pyroomacoustics

    Room Acoustics Simulation (RIR generator)

  • PyRubberband

    Wrapper for to do pitch-shifting and time-stretching

  • PyWavelets

    Discrete Wavelet Transform in Python

  • Resampy

    Sample rate conversion

  • SFS-Python

    Sound Field Synthesis Toolbox

  • sound_field_analysis

    Analyze, visualize and process sound field data recorded by spherical microphone arrays

  • STFT

    Standalone package for Short-Time Fourier Transform

  • aubio

    Feature extractor, written in C, Python interface

  • audioFlux

    A library for audio and music analysis, feature extraction

  • audiolazy

    Realtime Audio Processing lib, general purpose

  • essentia

    Music related low level and high level feature extractor, C++ based, includes Python bindings

  • python_speech_features

    Common speech features for ASR

  • pyYAAFE

    Python bindings for YAAFE feature extractor

  • speechpy

    Library for Speech Processing and Recognition, mostly feature extraction for now

  • spafe

    Python library for features extraction from audio files

  • audiomentations

    Audio Data Augmentation

  • muda

    Musical Data Augmentation

  • pydiogment

    Audio Data Augmentation

  • aeneas

    Forced aligner, based on MFCC+DTW, 35+ languages

  • deepspeech

    Pretrained automatic speech recognition

  • gentle

    Forced-aligner built on Kaldi

  • Parselmouth

    Python interface to the phonetics and speech analysis, synthesis, and manipulation software

  • persephone

    Automatic phoneme transcription tool

  • pyannote.audio

    Neural building blocks for speaker diarization

  • pyAudioAnalysis

    ² - Feature Extraction, Classification, Diarization

  • py-webrtcvad

    Interface to the WebRTC Voice Activity Detector

  • pypesq

    Wrapper for the PESQ score calculation

  • pystoi

    Short Term Objective Intelligibility measure (STOI)

  • PyWorldVocoder

    Wrapper for Morise's World Vocoder

  • Montreal Forced Aligner

    Forced aligner, based on Kaldi (HMM), English (others can be trained)

  • SIDEKIT

    Speaker and Language recognition

  • SpeechRecognition

    Wrapper for several ASR engines and APIs, online and offline

  • sed_eval

    Evaluation toolbox for Sound Event Detection

  • cochlea

    Inner ear models

  • Brian2

    Spiking neural networks simulator, includes cochlea model

  • Loudness

    Perceived loudness, includes Zwicker, Moore/Glasberg model

  • pyloudnorm

    Audio loudness meter and normalization, implements ITU-R BS.1770-4

  • Sound Field Synthesis Toolbox

    Sound Field Synthesis Toolbox

  • commonfate

    Common Fate Model and Transform

  • NTFLib

    Sparse Beta-Divergence Tensor Factorization

  • NUSSL

    Holistic source separation framework including DSP methods and deep learning methods

  • NIMFA

    Several flavors of non-negative-matrix factorization

  • Catchy

    Corpus Analysis Tools for Computational Hook Discovery

  • chord-detection

    Algorithms for chord detection and key estimation

  • Madmom

    MIR packages with strong focus on beat detection, onset detection and chord recognition

  • mir_eval

    Common scores for various MIR tasks. Also includes bss_eval implementation

  • msaf

    Music Structure Analysis Framework

  • librosa

    General audio and music analysis

  • Kapre

    Keras Audio Preprocessors

  • TorchAudio

    PyTorch Audio Loaders

  • nnAudio

    Accelerated audio processing using 1D convolution networks in PyTorch

  • Music21

    Toolkit for Computer-Aided Musicology

  • Mido

    Realtime MIDI wrapper

  • mingus

    Advanced music theory and notation package with MIDI file and playback support

  • Pretty-MIDI

    Utility functions for handling MIDI data in a nice/intuitive way

  • Jupylet

    Subtractive, additive, FM, and sample-based sound synthesis

  • PYO

    Realtime audio dsp engine

  • python-sounddevice

    PortAudio wrapper providing realtime audio I/O with NumPy

  • ReTiSAR

    Binarual rendering of streamed or IR-based high-order spherical microphone array signals

  • TimeSide (Beta)

    high level audio analysis, imaging, transcoding, streaming and labelling

  • beets

    Music library manager and tagger

  • musdb

    Parse and process the MUSDB18 dataset

  • medleydb

    Parse audio + annotations

  • Soundcloud API

    Wrapper for

  • Youtube-Downloader

    Download youtube videos (and the audio)

  • audiomate

    Loading different types of audio datasets

  • mirdata

    Common loaders for Music Information Retrieval (MIR) datasets

  • VamPy Host

    Interface compiled vamp plugins

Tutorials

Books

Scientific Papers

Other Resources

More related projects

Add a GitHub project

Missing a project or an awesome list? Paste its GitHub URL and we fetch it right away.