Back to GitHub trending

abus-aikorea/voice-pro

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

Python13K stars1.8K forks+58 today
View on GitHubShare on X

Topics

#faster-whisper#tts#whisper#gradio#subtitles#transcription#translator#webui#speech-recognition#speech-synthesis#speech-to-text#text-to-speech#yt-dlp#voice-cloning#podcasts#audiobook#voice-conversion#karaoke#whisperx

What it does

Voice-Pro is an advanced web application that integrates speech recognition, translation, and text-to-speech functionalities into a single tool for content creators. It supports multiple languages and offers voice cloning capabilities.

Star history

2026-08-23 → 2026-09-12+207 stars

Tracking

Last trending
2026-08-02

Creator kit

Hook

Transform your multimedia content creation with Voice-Pro, the ultimate AI-powered web app for speech recognition, translation, and dubbing. #AI #ContentCreation

Content angles

  • Create multilingual podcasts using Voice-Pro's text-to-speech and voice cloning features.
  • Demonstrate how to isolate vocals from YouTube videos using Demucs vocal isolation tool in Voice-Pro.
  • Showcase the process of translating and dubbing a video clip into multiple languages with Voice-Pro.

Who should care

Content creators, podcasters, researchers, developers interested in AI-driven multimedia solutions.