Open SourceVoice & Media11k
Vosk
Lightweight offline speech recognition toolkit for 20+ languages and embedded devices.
Vosk is an offline speech recognition toolkit with bindings for Python, Java, Node, C#, and more, and models for over twenty languages and dialects. Models are small (around 50 MB for the compact variants), making it well suited for mobile, Raspberry Pi, and on-device assistants while still supporting large server-grade models.
Repository
alphacep/vosk-apiLanguage
Jupyter NotebookWhat you'd build with it
- Embed offline voice command or dictation into a mobile or IoT product without cloud calls
- Stream microphone audio to a local recognizer with low latency for a voice assistant
- Run speaker-independent transcription in a language not well served by larger models
Tags
sttofflineembeddedmultilingual