music-audio-tagging-at-scale-models

Audio tagging research

Research on end-to-end learning for music audio tagging using large datasets and different front-end paradigms.

Tensorflow implementation of the models used in "End-to-end learning for music audio tagging at scale"

GitHub

149 stars
6 watching
19 forks
Language: Python
last commit: almost 7 years ago

Related projects:

RepositoryDescriptionStars
jordipons/eusipco2017Research code for music auto-tagging using deep learning and feature extraction23
jongpillee/musictagging_msdThis project is an audio classification system trained on the MSD tagging dataset, enabling automatic tagging of music files with relevant genres and styles.7
microsoft/pengiAn Audio Language Model framework that uses transfer learning to generate text from audio inputs295
balavenkatesh3322/audio-pretrained-modelA collection of pre-trained audio and speech models for various applications183
ibm/max-audio-classifierIdentifies sounds in short audio clips using machine learning and PCA transformation154
yuangongnd/whisper-atAn audio processing model that adds audio event tagging capabilities to an existing speech recognition system with minimal additional computational cost.343
jthorborg/apeAn Audio Programming Environment with support for AU and DSP plugins14
iver56/audiomentationsLibrary for audio data augmentation used in machine learning1,903
yuangongnd/ltuAn audio and speech large language model implementation with pre-trained models, datasets, and inference options396
kristijanbartol/deep-music-taggerClassifies music genres based on audio features using a deep learning model68
soundio/soundstageA graph object model and sequencing engine for Web Audio processing graphs65
soerenab/audiomnistThis project provides an implementation of a deep learning framework to classify audio signals and offers insights into the model's decision-making process using Explainable Artificial Intelligence (AI) techniques.351
ynop/audiomateA Python library for handling audio datasets, providing tools for accessing, manipulating, and preparing data for machine learning tasks.133
cpjku/madmomA Python audio signal processing library used in music information retrieval tasks.1,366
keunwoochoi/auralisationReconstructs audio features learned by convolutional neural networks into audible sounds42