vad

VAD system

An audio processing system that uses Deep Neural Networks and feature fusion to detect voice activity in speech recordings.

Voice Activity Detection system (Matlab-based implementation)

GitHub

105 stars
11 watching
48 forks
Language: Matlab
last commit: over 9 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
zhenghuatan/rvadAn unsupervised method for detecting speech activity in noisy audio signals130
yashdv/speech-recognitionA Matlab code to recognize individuals based on their unique vocal characteristics.40
wiseman/py-webrtcvadA Python interface to the WebRTC Voice Activity Detector2,088
matlab-deep-learning/wav2vec-2.0Enables speech-to-text transcription using a pre-trained neural network model in MATLAB.7
vocalpy/vakA Python framework for training and applying neural networks to acoustic communication research78
damo-nlp-sg/vcdAn approach to reduce object hallucinations in large vision-language models by contrasting output distributions derived from original and distorted visual inputs222
cvde/roomreverbA software tool for adding algorithmic reverb to audio recordings111
gsoh/vedA large-scale dataset capturing vehicle energy consumption and usage patterns in real-world driving scenarios.94
cvondrick/vaticTools for efficiently scaling up video annotation using crowdsourced marketplaces.609
vsitzmann/sirenAn implementation of a neural network architecture for implicit function representation learning using periodic activation functions.1,776
vehicle-lang/vehicleA toolkit for enforcing logical specifications on neural networks82
ahunnargikar/vagrant-mesosA Vagrant setup to create a Mesos/Docker/Marathon/Aurora/ Jenkins development environment for testing and development.122
marvinteichmann/multinetAn autonomous driving system that performs real-time road segmentation, car detection, and street classification using deep learning models.549
vadymmarkov/beethovenA Swift library providing an interface to pitch detection in audio signals.828
ksw0306/clarinetAn implementation of a neural network-based vocoder using parallel-wavenet architecture and autoregressive flow290