Open-source repositories gaining traction right now.
13 repositories
#speech-to-text
Sponsor Star cjpais / Handy A free, open source, and extensible speech-to-text application that works completely offline.

Star screenpipe / screenpipe screenpipe turns your computer into a personal AI that knows everything you've done. record. search. automate. all local, all private, all yours.

Build local voice agents with open-source models

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

Open Source Voice Agent Platform

ggml speech-to-text inference for 16+ model families

Local AI anywhere, for everyone — LLM inference, chat UI, voice, agents, workflows, RAG, and image generation. No cloud, no subscriptions.

Real-time speech-to-text caption appliance for a deaf user. Raspberry Pi + 10" touchscreen that transcribes phone calls and room conversation in near real-time.

Push-to-talk voice typing for your terminal. Local Whisper, cross-platform.

macOS voice productivity app — built-in dictation, AI rewrite, and translation. Powered by local Whisper + LLM.

Your voice, unheard by others. Privacy-first voice-to-text for macOS, built with Rust.

⚡ Offline speech-to-text plugin for OpenClaw — 1s cold start, 228MB model, Chinese-optimized. Powered by SenseVoice + Rust + ONNX Runtime. 离线语音转文字插件,秒级响应,中文优化,支持中/英/日/韩/粤五语。

Real-time voice-to-text input on Linux.
Paste a github.com URL. Submissions are reviewed before they are tracked.