jingjuSingingPhraseMatching

Audio-score matcher

This repository provides a software framework to match singing audio with corresponding music scores based on phonetic and duration information.

Code for the paper: Audio to Score Matching by Combining Phonetic and Duration Information

GitHub

27 stars
2 watching
6 forks
Language: Python
last commit: about 9 years ago
acoustic-modelcnn-modelhsmmphonemescoresinging-phrase

Related projects:

RepositoryDescriptionStars
ronggong/jingjusyllabicsegmentaionAn implementation of a score-informed method for segmenting jingju a cappella singing voice into syllables using convolutional neural networks and Viterbi algorithm7
sergree/matcheringAn audio matching and mastering tool that uses machine learning to adapt the sound of one track to match another1,842
ronggong/eusipco2017A software project that enables phoneme classification in music audio signals using convolutional neural networks and other machine learning techniques.19
ibm/max-chinese-phonetic-similarity-estimatorEstimates phonetic similarity between Chinese words and suggests similar-sounding candidates35
system-t/dimsimA phonetic similarity algorithm for indexing Chinese characters by sound123
bgutter/cl-phoneticProvides phonetic pattern matching functionality in Common Lisp to aid with natural language processing and text analysis.24
yuangongnd/ltuAn audio and speech large language model implementation with pre-trained models, datasets, and inference options396
jordipons/music-audio-tagging-at-scale-modelsResearch on end-to-end learning for music audio tagging using large datasets and different front-end paradigms.149
jamesturk/jellyfishA Python library providing algorithms and encoding schemes for approximate string matching.2,075
cpjku/madmomA Python audio signal processing library used in music information retrieval tasks.1,366
igglybuff/mregAn application that generates a string expression for filtering movie releases.15
xidongwu/d-auprcProvides an implementation of a specific algorithm used in audio signal processing0
iver56/audiomentationsLibrary for audio data augmentation used in machine learning1,903
yuangongnd/whisper-atAn audio processing model that adds audio event tagging capabilities to an existing speech recognition system with minimal additional computational cost.343
r3gm/sonitranslateSoftware that allows video translation with synchronized audio, utilizing speech-to-text and text-to-speech technologies.924