Open-source repositories gaining traction right now.
11 repositories
#speech-to-text
ggml speech-to-text inference for 16+ model families

Build local voice agents with open-source models

Open Source Voice Agent Platform

Local AI anywhere, for everyone — LLM inference, chat UI, voice, agents, workflows, RAG, and image generation. No cloud, no subscriptions.

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

Sponsor Star cjpais / Handy A free, open source, and extensible speech-to-text application that works completely offline.

Star screenpipe / screenpipe screenpipe turns your computer into a personal AI that knows everything you've done. record. search. automate. all local, all private, all yours.

Real-time speech-to-text caption appliance for a deaf user. Raspberry Pi + 10" touchscreen that transcribes phone calls and room conversation in near real-time.

Your voice, unheard by others. Privacy-first voice-to-text for macOS, built with Rust.

⚡ Offline speech-to-text plugin for OpenClaw — 1s cold start, 228MB model, Chinese-optimized. Powered by SenseVoice + Rust + ONNX Runtime. 离线语音转文字插件,秒级响应,中文优化,支持中/英/日/韩/粤五语。

Real-time voice-to-text input on Linux.
Paste a github.com URL. Submissions are reviewed before they are tracked.