STT

STT toolkit

A toolkit for building and deploying speech-to-text models using deep learning techniques

🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

GitHub

2k stars
62 watching
278 forks
Language: C++
last commit: over 2 years ago
Linked from 1 awesome list

asrautomatic-speech-recognitiondeep-learningspeech-recognitionspeech-recognition-apispeech-recognizerspeech-to-textstttensorflowvoice-recognition

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
coqui-ai/ttsA deep learning toolkit for generating human-like speech from text36,118
conchylicultor/deepqaA deep learning-based chatbot model using TensorFlow and RNNs to generate responses to user queries.2,929
rvc-boss/gpt-sovitsAn AI system for generating human-like voices from text inputs, using deep learning techniques and pre-trained models.36,977
openvinotoolkit/open_model_zooA collection of pre-trained deep learning models and demo applications for accelerating inference tasks4,118
tensorspeech/tensorflowttsReal-time speech synthesis using state-of-the-art architectures3,855
microsoft/deepspeedA deep learning optimization library that simplifies distributed training and inference on modern computing hardware.35,863
eleutherai/gpt-neoxProvides a framework for training large-scale language models on GPUs with advanced features and optimizations.6,997
deci-ai/super-gradientsA unified library for building and fine-tuning state-of-the-art computer vision models4,625
huggingface/transformersA collection of pre-trained machine learning models for various natural language and computer vision tasks, enabling developers to fine-tune and deploy these models on their own projects.136,357
replicate/cogA tool for packaging and deploying machine learning models in a standard, production-ready container environment.8,169
openvinotoolkit/openvinoA toolkit for optimizing and deploying artificial intelligence models in various applications7,439
dmlc/gluon-cvA toolkit for building and deploying deep learning models in computer vision5,850
jasonppy/voicecraftA neural codec model for speech editing and text-to-speech synthesis in real-time, using few seconds of reference audio.7,744
thudm/cogvlmDevelops a state-of-the-art visual language model with applications in image understanding and dialogue systems.6,182
mozilla/ttsAn open-source project providing a suite of deep learning models and tools for advanced text-to-speech synthesis.9,466