Open SourceVoice & Media2.5k
Coqui STT
Deep-learning speech-to-text engine that runs offline on a wide range of devices.
Coqui STT is an open-source speech-to-text engine descended from Mozilla DeepSpeech, designed to run fast enough for real-time use on devices from datacenters down to a Raspberry Pi. It ships pretrained models in many languages plus tooling to train and fine-tune custom acoustic models for domain-specific vocabularies.
Repository
coqui-ai/STTLanguage
C++What you'd build with it
- Add fully offline, privacy-preserving voice transcription to embedded or edge applications
- Fine-tune a custom STT model for a specialized domain vocabulary or accent
- Build a self-hosted dictation or command-recognition layer without a cloud API
Tags
sttofflinedeepspeechedge