@GitTrend0x: Holy cow, guys! Run voice cloning and cinematic video dubbing locally, supporting 646 languages, fully offline, no API key, no internet needed. ElevenLabs is crushed! https://github.com/debpalash/OmniVoice-Studio… This open-source marvel is insane...
Summary
OmniVoice Studio is an open-source desktop app that enables local voice cloning and cinematic video dubbing across 646 languages, fully offline with no API keys, positioning itself as a privacy-focused alternative to ElevenLabs.
View Cached Full Text
Cached at: 05/14/26, 02:29 AM
OmniVoice Studio
The open-source ElevenLabs alternative.
Real-time dictation, zero-shot voice cloning, and cinematic video dubbing — all on your desktop. Open-source, no API keys, fully local. 646 languages.
Quickstart · Features · Why OmniVoice Studio? · TTS Engines · Contributing · Discord
🎙️ Voice Cloning
3-second clip → mirror any voice. 646 languages, zero-shot.
🎨 Voice Design
Gender, age, accent, pitch, speed, emotion, dialect — dial it in.
🎬 Video Dubbing
YouTube URL or file → transcribe → translate → re-voice → MP4.
⌨️ Dictation Widget
⌘+⇧+Space from any app. Transcribes, auto-pastes, disappears.
🔊 Vocal Isolation
Demucs-powered. Splits speech from music, keeps the background.
👥 Speaker Diarization
Pyannote + WhisperX. Auto-identifies who said what.
📦 Batch Queue
Drop 50 videos, walk away. Progress bars per job.
🤖 MCP Server
Use OmniVoice from Claude, Cursor, or any MCP client.
🛡️ AI Watermark
AudioSeal (Meta). Invisible, survives compression.
🔐 100% Local
No keys, no cloud, no accounts. Your machine only.
⚡ GPU Auto-Detect
CUDA · MPS · ROCm · CPU. ≤8 GB? Auto-offloads.
🧩 Extensible
Subclass TTSBackend, add any engine in ~50 lines.
🖥️ Desktop App
🐳 Docker
⚡ From Source
Similar Articles
@cevenif: Bro, it's time to say goodbye to those paid voice tools! The open-source and free Voicebox has arrived, completely crushing paid giants like ElevenLabs and WisprFlow. Features: Voice cloning - instantly become anyone, Global voice input - accessible anytime...
An open-source, free local voice AI studio that supports voice cloning, voice generation, and global dictation. No API key required, runs entirely locally, and serves as a free alternative to ElevenLabs and WisprFlow.
@0x0SojalSec: Stop paying for ElevenLabs or cloud TTS. Free Clone voice in just 3 seconds fully Locally on Laptop Turns your docs int…
A free, fully local voice cloning and TTS tool powered by Qwen3-TTS and Kokoro runs on Apple Silicon via MLX, enabling studio-quality audiobook generation from PDFs without cloud services.
@Fluyeporlaweb: ElevenLabs costs $700 a year. HeyGen another $700. Someone just posted the local dubbing study that eliminates both sub…
OmniVoice Studio is a free, open-source tool that locally dubs MP4 videos into 600 languages using Whisper for transcription, voice cloning from 3 seconds of audio, and Demucs for background separation, eliminating the need for paid subscriptions like ElevenLabs and HeyGen.
@noahduck283: A tool that can download any YouTube video, cleanly remove vocals, transcribe, translate into 100+ languages, clone the original voice, and perform fully automatic dubbing. It takes less than 2 minutes. 100% runs locally. Free. Sews six top open-source models into a web page for "one-click download, vocal removal, transcription, translation, dubbing"...
Voice-Pro is a web tool that integrates six top open-source models (Whisper, Demucs, CosyVoice, F5-TTS, etc.), supporting YouTube video downloading, vocal removal, transcription, translation, voice cloning, and fully automatic dubbing. It takes less than 2 minutes, runs 100% locally, and is free.
@yiyirats: When relying on third-party voice services, data privacy and stability are always factors to consider. Voicebox is a locally run AI voice studio; all processing is done locally, data is not uploaded to the cloud, and no account registration is required. The features are quite comprehensive: supports Qwen3-TTS, LuxTTS, Chatterbo…
Voicebox is a locally run open-source AI voice studio that supports 7 TTS engines, 23 languages, and voice cloning. All processing is done locally to protect privacy. The project has received 33.8k stars on GitHub.