All open source
Open SourceVoice & Media2.5k

Coqui STT

Deep-learning speech-to-text engine that runs offline on a wide range of devices.

Coqui STT is an open-source speech-to-text engine descended from Mozilla DeepSpeech, designed to run fast enough for real-time use on devices from datacenters down to a Raspberry Pi. It ships pretrained models in many languages plus tooling to train and fine-tune custom acoustic models for domain-specific vocabularies.

Repository

coqui-ai/STT

Language

C++

What you'd build with it

  • Add fully offline, privacy-preserving voice transcription to embedded or edge applications
  • Fine-tune a custom STT model for a specialized domain vocabulary or accent
  • Build a self-hosted dictation or command-recognition layer without a cloud API

Tags

sttofflinedeepspeechedge