An open API service indexing awesome lists of open source software.

Projects in Awesome Lists tagged with vosk

A curated list of projects in awesome lists tagged with vosk .

https://github.com/alphacep/vosk-server

WebSocket, gRPC and WebRTC speech recognition server based on Vosk and Kaldi libraries

asr grpc kaldi python saas speech-recognition vosk webrtc websocket

Last synced: 11 Jun 2025

https://github.com/VocaHQ/vocalinux

Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!

accessibility dictation gpu-acceleration linux offline-first privacy-first python speech-recognition speech-to-text voice voice-typing vosk wayland whisper whisper-cpp

Last synced: 31 Aug 2026

https://github.com/ccoreilly/vosk-browser

A speech recognition library running in the browser thanks to a WebAssembly build of Vosk

asr kaldi speech-recognition speech-to-text stt typescript vosk wasm webassembly

Last synced: 16 May 2025

https://github.com/jatinkrmalik/vocalinux

Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!

accessibility dictation gpu-acceleration linux offline-first privacy-first python speech-recognition speech-to-text voice voice-typing vosk wayland whisper whisper-cpp

Last synced: 06 Jun 2026

https://github.com/ccoreilly/localstt

Android Speech Recognition Service using Vosk/Kaldi and Mozilla DeepSpeech

android deepspeech speech-recognition vosk

Last synced: 15 Apr 2025

https://github.com/riderodd/react-native-vosk

Speech recognition module for react native using Vosk library

asr react-native speech-recognition vosk

Last synced: 27 Feb 2026

https://github.com/papoteur-mga/elograf

Utility for launching and configuring nerd-dictation

nerd-dictation voice vosk

Last synced: 24 Feb 2026

https://github.com/sskorol/vosk-api-gpu

Vosk ASR Docker images with GPU for Jetson boards, PCs, M1 laptops and GPC

asr cuda docker gcp gpu jetson jetson-nano jetson-xavier-nx m1 nvidia nvidia-docker vosk vosk-api

Last synced: 23 Mar 2025

https://github.com/solaoi/lycoris

Real-time speech recognition & AI-powered note-taking app for macOS with offline/online modes, multilingual transcription, and Japanese translation support.

macos mcp mcp-client openai screenshot speech-recognition speech-to-text style-bert-vits2 voice-recognition vosk whisper

Last synced: 04 Apr 2026

https://github.com/botbahlul/vosk-powered-live-subtitle-v3

ANDROID APP that can RECOGNIZE ANY LIVE AUDIO/VIDEO STREAMING (using free VOSK Speech Recognition API) then TRANSLATE (using unofficial online Google Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE

android auto-caption auto-subtitle google-translate-api java live-subtitle speech-recognition voice-recognition vosk

Last synced: 11 Apr 2025

https://github.com/botbahlul/pyvosklivesubtitle

PySimpleGUI based DESKTOP APP that can RECOGNIZE any live streaming in 23 languages that supported by VOSK then TRANSLATE (using unofficial online Google Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE

auto-caption caption ffmpeg google-translate-api live-caption live-subtitle pysimplegui python speech-recognition speechrecognition subtitle voice-recognition voicerecognition vosk

Last synced: 27 Jul 2025

https://github.com/DanielTerletzkiy/chat-gtp-assistant

nodejs script that enables continuous conversation with voice recognition and tts speaker responses

googletts gptchat microphone nodejs speach-to-text speaker typescript vosk

Last synced: 24 Jul 2025

https://github.com/danielterletzkiy/chat-gtp-assistant

nodejs script that enables continuous conversation with voice recognition and tts speaker responses

googletts gptchat microphone nodejs speach-to-text speaker typescript vosk

Last synced: 03 Jul 2025

https://github.com/botbahlul/vosk_autosrt

A python script COMMAND LINE utility to AUTO GENERATE SUBTITLE FILE (using free Vosk Speech Recognition API) and TRANSLATED SUBTITLE FILE (using unofficial online Google Translate API) for any video or audio file

auto-caption auto-subtitle caption ffmpeg google-translate-api python speech-recognition speechrecognition subtitle voice-recognition voicerecognition vosk

Last synced: 28 Oct 2025

https://github.com/ocatias/AutoMash

Automatically create YouTube mashups. Given videos and a text, AutoMash will cut the videos together so the speakers in the video appears to says the given text.

creative-coding deep-speech deepspeech ibm-watson-speech meme-generator memes speech-recognition speech-to-text vosk youtube

Last synced: 09 Jul 2025

https://github.com/ccoreilly/catalan-speech-recognition-benchmark

A benchmark of speech recognition solutions for the Catalan language

asr asr-model catala catalan catalan-language deepspeech speech-recognition speech-to-text vosk

Last synced: 11 Jan 2026

https://github.com/smorodov/kaldi_vosk_win_cmake

cmake based kaldi + vosk + microphone speech recognition example

kaldi speaker-recognition speech-recognition speech-to-text voice-recognition vosk

Last synced: 10 Apr 2025

https://github.com/ggorets0dev/rantovox-telegram-bot

Telegram bot for text-to-speech and speech-to-speech translation, works with English and Russian languages.

console-application makefile pyttsx3 speech-recognition sst telegram-bot text-recognition tts vosk

Last synced: 01 Sep 2026

https://github.com/msqr1/vosklet

A speech recognizer that can run on the browser, inspired by vosk-browser

api javascript speech-recognizer vosk webassembly

Last synced: 25 Sep 2025

https://github.com/sskorol/respeaker-websockets

This project reveals full Respeaker Core V2 potential by using bundled Alango DSP algorithms for audio pre-processing before sending to the custom ASR server.

alango asr dsp pixel-ring respeaker respeaker-core-v2 vosk websockets

Last synced: 15 May 2025

https://github.com/ave-sergeev/dictator

Speech-to-Text translation service (Rust, Tonic) (2025)

audio rust silero tonic transcribe voice-activity-detection vosk

Last synced: 03 Apr 2025

https://github.com/puff-dayo/wxreader

A lightweight PDF, ePub and ZIP book/manga reader app for Windows/Linux. Custom OpenGL shaders, external control (voice and webcam eye-gesture) supported.

book-reader comics-reader desktop-app eye-tracking gui-application linux-app pdf-viewer pymupdf pyopengl pyvips shader-effects vosk windows-app wxpython

Last synced: 26 Apr 2026

https://github.com/ascender1729/audiodictate

An efficient desktop application for transcribing audio files into text using Vosk speech recognition.

audio-processing audio-to-text offline-transcription python speech-recognition transcription vosk

Last synced: 18 Oct 2025

https://github.com/botbahlul/java-vosk-livesubtitle

JAVA based DESKTOP APP that can RECOGNIZE any live streaming in 21 languages that supported by VOSK then TRANSLATE (using unofficial online Google Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE

auto-caption auto-subtitle google-translate-api java live-caption live-subtitle speech-recognition voice-recognition vosk

Last synced: 11 Apr 2025

https://github.com/thunderpoot/audio-censor

Rudimentary program for speech transcription, manipulation, and redaction.

audio censor censorship pydub redaction speech transcription vosk vosk-models wavesurfer

Last synced: 27 May 2026

https://github.com/hwpoison/vosk-voice-recognition-c

Offline voice recognition using pure C and vosk lib. (from file and from microphone, windows)

c speech-recognition speech-to-text voice-recognition vosk vosk-api vosk-models wav windows winmm

Last synced: 12 Aug 2025

https://github.com/menshovanton/voiceassistantleo

Лео — голосовой помощник для Windows. Написанный на C#. Четко распознает голос с помощью Vosk STT. Озвучен при помощи Silero TTS.

csharp dotnet gui oss russian-language sileros-tts voice-assistant voice-commands voice-recognition vosk vosk-engine wakeword windows-10 windows-11 windows-desktop wpf

Last synced: 22 Jul 2025

https://github.com/tez3998/audio-output-to-text

VOSKを使ったスピーカーやヘッドフォンから出力される音声のオフライン文字起こし

headphones speaker speech-recognition vosk

Last synced: 01 Aug 2025

https://github.com/anuran-roy/vosk-demo

A simple offline voice recognition system purely built on Python3, that prints the keywords on the terminal screen. It uses Vosk as the backend, and NLTK to extract keywords from the text generated.

offline-capable python3 speech-recognition speech-to-text vosk

Last synced: 20 May 2026

https://github.com/botbahlul/vosk-powered-live-subtitle

ANDROID APP that can RECOGNIZE ANY LIVE AUDIO/VIDEO STREAMING (using VOSK Speech Recognition API) then TRANSLATE (using ANDROID MLKIT TRANSLATE API) and display it as LIVE CAPTION / LIVE SUBTITLE

android auto-caption auto-subtitle java live-subtitle mlkit-translate speech-recognition voice-recognition vosk

Last synced: 11 Apr 2025

https://github.com/hwpoison/voice-assistant

an offline small voice desktop assistant written on a booring sunday afternoon, uses autoit for automatitation, vosk for voice recognition a gtts for speech synth.

assistant autoit gtts linux voice-recognition-experiment vosk windows

Last synced: 09 Mar 2026

https://github.com/zahraarshia/speech-summarization

Persian Speech Summarization Dataset and Test Bench Model

datase nlp persian speech speech-summarizer torch vosk

Last synced: 21 Aug 2025

https://github.com/meganetaaan/simple-stt-server

A simple text-to-speech server that uses VOSK to recognize speech and send it over WebSocket

nodejs speech-recognition speech-to-text vosk

Last synced: 31 Jan 2026

https://github.com/jacoblincool/vosk-cli

Use vosk in command line. List all pre-trained models, download & install them, and use them to transcribe audio files or live audio.

speech-recognition vosk

Last synced: 01 Jul 2025

https://github.com/botbahlul/vosk-powered-live-subtitle-v2

ANDROID APP that can RECOGNIZE LIVE AUDIO/VIDEO STREAMING (using free VOSK Speech Recognition API) then TRANSLATE (using MLKIT Translate API) and display it as LIVE CAPTION / LIVE SUBTITLE

android auto-caption auto-subtitle java live-subtitle mlkit-translate speech-recognition speech-to-text voice-recognition vosk

Last synced: 22 Sep 2025

https://github.com/MenshovAnton/VoiceAssistantLeo

Лео — голосовой помощник для Windows. Написанный на C#. Четко распознает голос с помощью Vosk STT. Озвучен при помощи Silero TTS.

csharp dotnet gui oss russian-language sileros-tts voice-assistant voice-commands voice-recognition vosk vosk-engine wakeword windows-10 windows-11 windows-desktop wpf

Last synced: 17 Jul 2025

https://github.com/maniam/speak-io

A web API for speech-to-text (STT) and text-to-speech (TTS) that integrates with existing engines, supporting real-time audio streaming and modular engine selection.

bark-tts chatterbox-tts conqui-tts fast-whisper fastapi piper-tts vosk websocket whisper-ai whisper-cpp

Last synced: 29 Apr 2026

https://github.com/lhg96/stt-demo-korean

Korean Speech-to-Text app with Whisper & Vosk | 한국어 음성인식 데모 애플리케이션

korean-stt real-time-audio speech-recognition speech-to-text vosk whisper

Last synced: 20 Jun 2026

https://github.com/subuhana2303/vaanirakshak_offline-emergency-voice-assistant

VaaniRakshak is an offline voice assistant built for disaster scenarios, enabling hands-free emergency support without internet connectivity. It assists users in locating shelters, requesting help, and accessing life-saving information through voice interaction.

audio-input json pyaudio pyttsx3 speech-recognition text-to-speech tkinter-gui vosk

Last synced: 23 Jul 2025

https://github.com/sskorol/matrix-voice-esp32-ws-streamer

Matrix Voice streaming via WebSockets.

asr esp32 matrix-voice speech-recognition vosk

Last synced: 15 May 2026

https://github.com/lasithaamarasinghe/real-time-speech-recognition

This project includes a system that can record live speech using your microphone and then transcribe it using speech recognition.

ipywidgets jupyter-notebook machine-learning pyaudio pydub python3 pytorch realtime-speech-recognition speech-to-text transformers vosk

Last synced: 06 May 2026

https://github.com/ryo08271154/voice_control

音声操作でスマートデバイスを操作できるPython製アプリケーションです。

chromecast flet mcp mcp-client python raspberry-pi raspberrypi smart-home smarthome switchbot-api voice-assistant voice-control vosk

Last synced: 16 Aug 2025

https://github.com/karun-16/privy-offline-voice-assistant

Privacy-first offline desktop voice assistant using Vosk speech recognition

desktop-assistant offline-ai privacy-first python speech-recognition voice-assistant vosk

Last synced: 20 Feb 2026

https://github.com/scaledteam/nerd-dictation-uinput

Simple speech to text using Vosk and Uinput with russian language support

linux russian russian-language uinput vosk

Last synced: 17 Apr 2026

https://github.com/m15-ai/faster-local-voice-ai

A real-time, fully local voice AI system optimized for low-resource devices like an 8GB Ubuntu laptop with no GPU, achieving sub-second STT-to-TTS latency using Ollama, Vosk, Piper, and JACK/PipeWire. Open-source and privacy-focused for offline conversational AI.

conversational-ai edge-ai jack local-ai low-latency offline-ai ollama piper pipewire python stt tts voice-ai vosk websockets

Last synced: 16 May 2026

https://github.com/karvy-singh/voice_controlled_-system_lock

Appear cool 😎, by locking your system with command similar to " % Lock %"

python sounddevice voice-commands vosk wave

Last synced: 29 Mar 2025

https://github.com/maniam/trigger-talk

Trigger-Talk is an offline-capable hotword detection engine that passively listens for custom wake phrases to trigger speech recognition or automation workflows.

hotword-detection openwakeword porcupine vosk wake-word-detection

Last synced: 01 Jul 2025

https://github.com/liquid36/whatsapp-audio-bot

Whasapp Speach to Text Bot

bot javascript node venon vosk whastapp

Last synced: 11 Apr 2026

https://github.com/jhaabhijeet864/wake_bot

a complex, multithreaded Python Windows application that is like an autonomous startup buddy for your system

audio python spotify vosk vscode windows

Last synced: 18 May 2026

https://github.com/rudradave1/sotto-app

Private voice journal for Android. Speak your thoughts, AI structures them, everything stays encrypted on your device.

android encrypted groq jetpack-compose koin kotlin kotlin-multiplatform ktor offline-first privacy sqlcipher sqldelight voice-journal vosk workmanager

Last synced: 02 Jul 2026

https://github.com/victor141516/vosk-http-api

I took Vosk and wrapped it in a HTTP API

api http vosk vosk-api

Last synced: 24 Mar 2025

https://github.com/prakash-aryan/speech_command_server

A real-time voice command detection system that recognizes "play" and "pause" commands using Vosk speech recognition.

fastapi python transcription uvicorn vosk websockets

Last synced: 17 May 2026

https://github.com/elara-mendes/hear-my-name

An app to help people with hearing disabilities.

disability-support python vosk

Last synced: 25 Apr 2025

https://github.com/seapagan/vosk-test

Some experiments in using 'Vosk' speech-to-text under Python including real-time from a microphone over the web.

python speech-to-text vosk vosk-api websocket

Last synced: 18 Jul 2026

https://github.com/rohitmalwal/ai_assistant

AI Computer Assistant: A Python-based virtual assistant that performs tasks like voice-controlled web automation, AI chat, Wikipedia search, weather updates, time announcements, etc.

ast opeanai-api openai python python3 pywin32 requests rich subprocess time vosk vosk-models warnings weather-api webbrowser wikipedia

Last synced: 14 Apr 2026

https://github.com/xuejiazhi/voice

语音识别模型

ai funasr vosk vosk-api vosk-models

Last synced: 31 Jan 2026

https://github.com/palaashatri/jvosk

Audio transcription using Vosk. Built with Swing.

gui java speech-recognition speech-to-text swing transcription vosk

Last synced: 01 Mar 2026

https://github.com/nougatcat/voice_assistant

[2022 AI] Интеллектуальный голосовой помощник на Python

python sklearn vosk

Last synced: 28 Apr 2026

https://github.com/borisboc/speech-markdown

Get the speech from audio (or video), and turn this into markdown. Mainly designed for serious / educational purposes

audio-extraction ollama ollama-api speech-to-text surfacing vosk vosk-api

Last synced: 23 Jun 2026