Projects in Awesome Lists tagged with vosk
A curated list of projects in awesome lists tagged with vosk .
https://github.com/alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
android asr deep-learning deep-neural-networks deepspeech google-speech-to-text ios kaldi offline privacy python raspberry-pi speaker-identification speaker-verification speech-recognition speech-to-text speech-to-text-android stt voice-recognition vosk
Last synced: 12 May 2025
https://github.com/stypox/dicio-android
Dicio assistant app for Android
android assistant assistive-technology dicio dicio-assistant personal-assistant personal-assistant-framework voice-assistant vosk
Last synced: 14 May 2025
https://github.com/Stypox/dicio-android
Dicio assistant app for Android
android assistant assistive-technology dicio dicio-assistant personal-assistant personal-assistant-framework voice-assistant vosk
Last synced: 05 Apr 2025
https://github.com/alphacep/vosk-server
WebSocket, gRPC and WebRTC speech recognition server based on Vosk and Kaldi libraries
asr grpc kaldi python saas speech-recognition vosk webrtc websocket
Last synced: 11 Jun 2025
https://github.com/VocaHQ/vocalinux
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
accessibility dictation gpu-acceleration linux offline-first privacy-first python speech-recognition speech-to-text voice voice-typing vosk wayland whisper whisper-cpp
Last synced: 31 Aug 2026
https://github.com/ccoreilly/vosk-browser
A speech recognition library running in the browser thanks to a WebAssembly build of Vosk
asr kaldi speech-recognition speech-to-text stt typescript vosk wasm webassembly
Last synced: 16 May 2025
https://github.com/jatinkrmalik/vocalinux
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
accessibility dictation gpu-acceleration linux offline-first privacy-first python speech-recognition speech-to-text voice voice-typing vosk wayland whisper whisper-cpp
Last synced: 06 Jun 2026
https://github.com/ccoreilly/localstt
Android Speech Recognition Service using Vosk/Kaldi and Mozilla DeepSpeech
android deepspeech speech-recognition vosk
Last synced: 15 Apr 2025
https://github.com/riderodd/react-native-vosk
Speech recognition module for react native using Vosk library
asr react-native speech-recognition vosk
Last synced: 27 Feb 2026
https://github.com/matteo-convertino/vosk-build-model
How to create your own model for vosk
deep-learning deep-neural-networks guide kaldi speech-recognition tutorial voice-recognition vosk walkthrough
Last synced: 10 Apr 2025
https://github.com/papoteur-mga/elograf
Utility for launching and configuring nerd-dictation
Last synced: 24 Feb 2026
https://github.com/sskorol/vosk-api-gpu
Vosk ASR Docker images with GPU for Jetson boards, PCs, M1 laptops and GPC
asr cuda docker gcp gpu jetson jetson-nano jetson-xavier-nx m1 nvidia nvidia-docker vosk vosk-api
Last synced: 23 Mar 2025
https://github.com/solaoi/lycoris
Real-time speech recognition & AI-powered note-taking app for macOS with offline/online modes, multilingual transcription, and Japanese translation support.
macos mcp mcp-client openai screenshot speech-recognition speech-to-text style-bert-vits2 voice-recognition vosk whisper
Last synced: 04 Apr 2026
https://github.com/mmguero/monkeyplug
monkeyplug is a little script to mute profanity in audio files
audio audio-files ffmpeg objectional-language obscene-filter obscenity podcasts profanity profanity-detection profanity-filter profanity-filtering profanityfilter python python3 swear-filter vosk
Last synced: 27 Jan 2026
https://github.com/botbahlul/vosk-powered-live-subtitle-v3
ANDROID APP that can RECOGNIZE ANY LIVE AUDIO/VIDEO STREAMING (using free VOSK Speech Recognition API) then TRANSLATE (using unofficial online Google Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE
android auto-caption auto-subtitle google-translate-api java live-subtitle speech-recognition voice-recognition vosk
Last synced: 11 Apr 2025
https://github.com/botbahlul/pyvosklivesubtitle
PySimpleGUI based DESKTOP APP that can RECOGNIZE any live streaming in 23 languages that supported by VOSK then TRANSLATE (using unofficial online Google Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE
auto-caption caption ffmpeg google-translate-api live-caption live-subtitle pysimplegui python speech-recognition speechrecognition subtitle voice-recognition voicerecognition vosk
Last synced: 27 Jul 2025
https://github.com/DanielTerletzkiy/chat-gtp-assistant
nodejs script that enables continuous conversation with voice recognition and tts speaker responses
googletts gptchat microphone nodejs speach-to-text speaker typescript vosk
Last synced: 24 Jul 2025
https://github.com/danielterletzkiy/chat-gtp-assistant
nodejs script that enables continuous conversation with voice recognition and tts speaker responses
googletts gptchat microphone nodejs speach-to-text speaker typescript vosk
Last synced: 03 Jul 2025
https://github.com/openvoiceos/ovos-stt-plugin-vosk
vosk STT plugin for mycroft
asr automatic-speech-recognition hacktoberfest kaldi speech-recognition speech-to-text stt vosk
Last synced: 16 May 2025
https://github.com/OpenVoiceOS/ovos-stt-plugin-vosk
vosk STT plugin for mycroft
asr automatic-speech-recognition hacktoberfest kaldi speech-recognition speech-to-text stt vosk
Last synced: 10 May 2025
https://github.com/botbahlul/vosk_autosrt
A python script COMMAND LINE utility to AUTO GENERATE SUBTITLE FILE (using free Vosk Speech Recognition API) and TRANSLATED SUBTITLE FILE (using unofficial online Google Translate API) for any video or audio file
auto-caption auto-subtitle caption ffmpeg google-translate-api python speech-recognition speechrecognition subtitle voice-recognition voicerecognition vosk
Last synced: 28 Oct 2025
https://github.com/ocatias/AutoMash
Automatically create YouTube mashups. Given videos and a text, AutoMash will cut the videos together so the speakers in the video appears to says the given text.
creative-coding deep-speech deepspeech ibm-watson-speech meme-generator memes speech-recognition speech-to-text vosk youtube
Last synced: 09 Jul 2025
https://github.com/ccoreilly/catalan-speech-recognition-benchmark
A benchmark of speech recognition solutions for the Catalan language
asr asr-model catala catalan catalan-language deepspeech speech-recognition speech-to-text vosk
Last synced: 11 Jan 2026
https://github.com/smorodov/kaldi_vosk_win_cmake
cmake based kaldi + vosk + microphone speech recognition example
kaldi speaker-recognition speech-recognition speech-to-text voice-recognition vosk
Last synced: 10 Apr 2025
https://github.com/ggorets0dev/rantovox-telegram-bot
Telegram bot for text-to-speech and speech-to-speech translation, works with English and Russian languages.
console-application makefile pyttsx3 speech-recognition sst telegram-bot text-recognition tts vosk
Last synced: 01 Sep 2026
https://github.com/msqr1/vosklet
A speech recognizer that can run on the browser, inspired by vosk-browser
api javascript speech-recognizer vosk webassembly
Last synced: 25 Sep 2025
https://github.com/sskorol/respeaker-websockets
This project reveals full Respeaker Core V2 potential by using bundled Alango DSP algorithms for audio pre-processing before sending to the custom ASR server.
alango asr dsp pixel-ring respeaker respeaker-core-v2 vosk websockets
Last synced: 15 May 2025
https://github.com/ave-sergeev/dictator
Speech-to-Text translation service (Rust, Tonic) (2025)
audio rust silero tonic transcribe voice-activity-detection vosk
Last synced: 03 Apr 2025
https://github.com/puff-dayo/wxreader
A lightweight PDF, ePub and ZIP book/manga reader app for Windows/Linux. Custom OpenGL shaders, external control (voice and webcam eye-gesture) supported.
book-reader comics-reader desktop-app eye-tracking gui-application linux-app pdf-viewer pymupdf pyopengl pyvips shader-effects vosk windows-app wxpython
Last synced: 26 Apr 2026
https://github.com/VoXera/VoXera
An Open-Source Persian Language Techs Toolkit with Python
deep-learning deep-neural-networks keyword-extraction machine-learning natural-language-processing nlp openai persian persian-language speech-recognition speech-to-text text-processing vosk vosk-api whisper
Last synced: 08 Jul 2025
https://github.com/ascender1729/audiodictate
An efficient desktop application for transcribing audio files into text using Vosk speech recognition.
audio-processing audio-to-text offline-transcription python speech-recognition transcription vosk
Last synced: 18 Oct 2025
https://github.com/takeoutfm/takeout_assistant
Offline voice assistant for Android
agplv3 android bloc clock flutter home-assistant hue hue-bridge hue-lights music speech-recognition voice-assistant vosk
Last synced: 27 Apr 2026
https://github.com/botbahlul/java-vosk-livesubtitle
JAVA based DESKTOP APP that can RECOGNIZE any live streaming in 21 languages that supported by VOSK then TRANSLATE (using unofficial online Google Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE
auto-caption auto-subtitle google-translate-api java live-caption live-subtitle speech-recognition voice-recognition vosk
Last synced: 11 Apr 2025
https://github.com/thunderpoot/audio-censor
Rudimentary program for speech transcription, manipulation, and redaction.
audio censor censorship pydub redaction speech transcription vosk vosk-models wavesurfer
Last synced: 27 May 2026
https://github.com/hwpoison/vosk-voice-recognition-c
Offline voice recognition using pure C and vosk lib. (from file and from microphone, windows)
c speech-recognition speech-to-text voice-recognition vosk vosk-api vosk-models wav windows winmm
Last synced: 12 Aug 2025
https://github.com/menshovanton/voiceassistantleo
Лео — голосовой помощник для Windows. Написанный на C#. Четко распознает голос с помощью Vosk STT. Озвучен при помощи Silero TTS.
csharp dotnet gui oss russian-language sileros-tts voice-assistant voice-commands voice-recognition vosk vosk-engine wakeword windows-10 windows-11 windows-desktop wpf
Last synced: 22 Jul 2025
https://github.com/tez3998/audio-output-to-text
VOSKを使ったスピーカーやヘッドフォンから出力される音声のオフライン文字起こし
headphones speaker speech-recognition vosk
Last synced: 01 Aug 2025
https://github.com/anuran-roy/vosk-demo
A simple offline voice recognition system purely built on Python3, that prints the keywords on the terminal screen. It uses Vosk as the backend, and NLTK to extract keywords from the text generated.
offline-capable python3 speech-recognition speech-to-text vosk
Last synced: 20 May 2026
https://github.com/botbahlul/vosk-powered-live-subtitle
ANDROID APP that can RECOGNIZE ANY LIVE AUDIO/VIDEO STREAMING (using VOSK Speech Recognition API) then TRANSLATE (using ANDROID MLKIT TRANSLATE API) and display it as LIVE CAPTION / LIVE SUBTITLE
android auto-caption auto-subtitle java live-subtitle mlkit-translate speech-recognition voice-recognition vosk
Last synced: 11 Apr 2025
https://github.com/hwpoison/voice-assistant
an offline small voice desktop assistant written on a booring sunday afternoon, uses autoit for automatitation, vosk for voice recognition a gtts for speech synth.
assistant autoit gtts linux voice-recognition-experiment vosk windows
Last synced: 09 Mar 2026
https://github.com/zahraarshia/speech-summarization
Persian Speech Summarization Dataset and Test Bench Model
datase nlp persian speech speech-summarizer torch vosk
Last synced: 21 Aug 2025
https://github.com/mycore-org/myvidcore
My Video Converter
convert-videos ffmpeg gpu-support gui java mycore vosk webservice
Last synced: 05 Sep 2025
https://github.com/meganetaaan/simple-stt-server
A simple text-to-speech server that uses VOSK to recognize speech and send it over WebSocket
nodejs speech-recognition speech-to-text vosk
Last synced: 31 Jan 2026
https://github.com/jacoblincool/vosk-cli
Use vosk in command line. List all pre-trained models, download & install them, and use them to transcribe audio files or live audio.
Last synced: 01 Jul 2025
https://github.com/ankushrathour/audio-visualization-and-speech-recognition
Convert audio to text using JavaScript, Speech To Text.
audio-player audio-visualizer javascript speech-recognition speech-to-text vosk wavesurfer-js
Last synced: 27 Nov 2025
https://github.com/botbahlul/vosk-powered-live-subtitle-v2
ANDROID APP that can RECOGNIZE LIVE AUDIO/VIDEO STREAMING (using free VOSK Speech Recognition API) then TRANSLATE (using MLKIT Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE
android auto-caption auto-subtitle java live-subtitle mlkit-translate speech-recognition speech-to-text voice-recognition vosk
Last synced: 22 Sep 2025
https://github.com/MenshovAnton/VoiceAssistantLeo
Лео — голосовой помощник для Windows. Написанный на C#. Четко распознает голос с помощью Vosk STT. Озвучен при помощи Silero TTS.
csharp dotnet gui oss russian-language sileros-tts voice-assistant voice-commands voice-recognition vosk vosk-engine wakeword windows-10 windows-11 windows-desktop wpf
Last synced: 17 Jul 2025
https://github.com/maniam/speak-io
A web API for speech-to-text (STT) and text-to-speech (TTS) that integrates with existing engines, supporting real-time audio streaming and modular engine selection.
bark-tts chatterbox-tts conqui-tts fast-whisper fastapi piper-tts vosk websocket whisper-ai whisper-cpp
Last synced: 29 Apr 2026
https://github.com/lhg96/stt-demo-korean
Korean Speech-to-Text app with Whisper & Vosk | 한국어 음성인식 데모 애플리케이션
korean-stt real-time-audio speech-recognition speech-to-text vosk whisper
Last synced: 20 Jun 2026
https://github.com/subuhana2303/vaanirakshak_offline-emergency-voice-assistant
VaaniRakshak is an offline voice assistant built for disaster scenarios, enabling hands-free emergency support without internet connectivity. It assists users in locating shelters, requesting help, and accessing life-saving information through voice interaction.
audio-input json pyaudio pyttsx3 speech-recognition text-to-speech tkinter-gui vosk
Last synced: 23 Jul 2025
https://github.com/sskorol/matrix-voice-esp32-ws-streamer
Matrix Voice streaming via WebSockets.
asr esp32 matrix-voice speech-recognition vosk
Last synced: 15 May 2026
https://github.com/lasithaamarasinghe/real-time-speech-recognition
This project includes a system that can record live speech using your microphone and then transcribe it using speech recognition.
ipywidgets jupyter-notebook machine-learning pyaudio pydub python3 pytorch realtime-speech-recognition speech-to-text transformers vosk
Last synced: 06 May 2026
https://github.com/cjh0613/vosk-android-demo-chinese
中文 vosk-android-demo
android asr offline offline-asr offline-speech-recognition speech-recognition vosk vosk-models
Last synced: 17 Jan 2026
https://github.com/ryo08271154/voice_control
音声操作でスマートデバイスを操作できるPython製アプリケーションです。
chromecast flet mcp mcp-client python raspberry-pi raspberrypi smart-home smarthome switchbot-api voice-assistant voice-control vosk
Last synced: 16 Aug 2025
https://github.com/karun-16/privy-offline-voice-assistant
Privacy-first offline desktop voice assistant using Vosk speech recognition
desktop-assistant offline-ai privacy-first python speech-recognition voice-assistant vosk
Last synced: 20 Feb 2026
https://github.com/scaledteam/nerd-dictation-uinput
Simple speech to text using Vosk and Uinput with russian language support
linux russian russian-language uinput vosk
Last synced: 17 Apr 2026
https://github.com/m15-ai/faster-local-voice-ai
A real-time, fully local voice AI system optimized for low-resource devices like an 8GB Ubuntu laptop with no GPU, achieving sub-second STT-to-TTS latency using Ollama, Vosk, Piper, and JACK/PipeWire. Open-source and privacy-focused for offline conversational AI.
conversational-ai edge-ai jack local-ai low-latency offline-ai ollama piper pipewire python stt tts voice-ai vosk websockets
Last synced: 16 May 2026
https://github.com/karvy-singh/voice_controlled_-system_lock
Appear cool 😎, by locking your system with command similar to " % Lock %"
python sounddevice voice-commands vosk wave
Last synced: 29 Mar 2025
https://github.com/maniam/trigger-talk
Trigger-Talk is an offline-capable hotword detection engine that passively listens for custom wake phrases to trigger speech recognition or automation workflows.
hotword-detection openwakeword porcupine vosk wake-word-detection
Last synced: 01 Jul 2025
https://github.com/liquid36/whatsapp-audio-bot
Whasapp Speach to Text Bot
bot javascript node venon vosk whastapp
Last synced: 11 Apr 2026
https://github.com/rudradave1/sotto-app
Private voice journal for Android. Speak your thoughts, AI structures them, everything stays encrypted on your device.
android encrypted groq jetpack-compose koin kotlin kotlin-multiplatform ktor offline-first privacy sqlcipher sqldelight voice-journal vosk workmanager
Last synced: 02 Jul 2026
https://github.com/victor141516/vosk-http-api
I took Vosk and wrapped it in a HTTP API
Last synced: 24 Mar 2025
https://github.com/prakash-aryan/speech_command_server
A real-time voice command detection system that recognizes "play" and "pause" commands using Vosk speech recognition.
fastapi python transcription uvicorn vosk websockets
Last synced: 17 May 2026
https://github.com/elara-mendes/hear-my-name
An app to help people with hearing disabilities.
disability-support python vosk
Last synced: 25 Apr 2025
https://github.com/a-iceberg/whisper_model_evaluator
WER, MER, WIL of Whisper vs Vosk vs Google transcribators comparator
asr audio-to-text automatic-speech-recognition data-analysis evaluation google-speech-recognition python tuning-parameters visualization vosk whisper
Last synced: 11 Mar 2025
https://github.com/seapagan/vosk-test
Some experiments in using 'Vosk' speech-to-text under Python including real-time from a microphone over the web.
python speech-to-text vosk vosk-api websocket
Last synced: 18 Jul 2026
https://github.com/rohitmalwal/ai_assistant
AI Computer Assistant: A Python-based virtual assistant that performs tasks like voice-controlled web automation, AI chat, Wikipedia search, weather updates, time announcements, etc.
ast opeanai-api openai python python3 pywin32 requests rich subprocess time vosk vosk-models warnings weather-api webbrowser wikipedia
Last synced: 14 Apr 2026
https://github.com/xuejiazhi/voice
语音识别模型
ai funasr vosk vosk-api vosk-models
Last synced: 31 Jan 2026
https://github.com/palaashatri/jvosk
Audio transcription using Vosk. Built with Swing.
gui java speech-recognition speech-to-text swing transcription vosk
Last synced: 01 Mar 2026
https://github.com/nougatcat/voice_assistant
[2022 AI] Интеллектуальный голосовой помощник на Python
Last synced: 28 Apr 2026
https://github.com/borisboc/speech-markdown
Get the speech from audio (or video), and turn this into markdown. Mainly designed for serious / educational purposes
audio-extraction ollama ollama-api speech-to-text surfacing vosk vosk-api
Last synced: 23 Jun 2026