@noahduck283: A tool that can download any YouTube video, cleanly remove vocals, transcribe, translate into 100+ languages, clone the original voice, and perform fully automatic dubbing. It takes less than 2 minutes. 100% runs locally. Free. Sews six top open-source models into a web page for "one-click download, vocal removal, transcription, translation, dubbing"...
Summary
Voice-Pro is a web tool that integrates six top open-source models (Whisper, Demucs, CosyVoice, F5-TTS, etc.), supporting YouTube video downloading, vocal removal, transcription, translation, voice cloning, and fully automatic dubbing. It takes less than 2 minutes, runs 100% locally, and is free.
View Cached Full Text
Cached at: 05/22/26, 11:59 PM
Voice-Pro
The best AI speech recognition, translation, and multilingual dubbing solution 🚀
Korean ∙ English ∙ Simplified Chinese ∙ Traditional Chinese ∙ Japanese ∙ German ∙ Spanish ∙ Portuguese
Similar Articles
@yhslgg: Old Yang shares another gem open-source tool—KrillinAI, 10,000 stars on GitHub, a must-see for multilingual audio/video content! In a nutshell: from video download to subtitle translation, AI dubbing, video compositing, the entire pipeline is covered, and it can even auto-generate platform covers, supporting Bilibili, Douyin, Xiaohongshu, YouTube…
KrillinAI is an open-source tool that integrates the entire workflow of video downloading, subtitle translation, AI dubbing, and video compositing. It supports context-aware translation, voice cloning, auto layout, and cover generation, and is compatible with multiple AI models, suitable for multilingual audio/video content creation and distribution.
@Huahuazo: After trying this voice cloning tool, I was stunned. I casually took an 8-second recording of myself speaking, threw it in, and the output voice—the tone, pauses, that cheeky tone—was exactly like me arguing. The best part is it runs completely locally, no audio upload needed, privacy and security are fully maximized. Honestly, my first reaction wasn't 'impressive,' but rather 'chills down my spine': in the future, just by a voice recording, you really can't take things too seriously…
Introducing the local voice cloning tool VoiceStudio, which only requires 8 seconds of recording to highly replicate the voice, runs completely locally to protect privacy, and supports multiple languages and platforms.
@FakeMaidenMaker: Explosive! This open-source project converts text to human-like voice for free, can clone anyone's voice, and adjust timbre with text! GitHub has garnered 30K stars, from Mianbao Intelligent OpenBMB, VoxCPM previously topped both GitHub and HuggingFace charts. Do...
VoxCPM2 is an open-source speech synthesis model from OpenBMB, using a tokenizer-free diffusion autoregressive architecture, supporting 30 languages, voice design, and controllable voice cloning. It can clone a voice with just one sentence, or create a brand new voice using text, outputting 48kHz high-quality audio, and is commercially usable.
@hank_aibtc: Whoa, this thing really blew my mind. Dug up a local tool called KrillinAI, free, video translation, precise subtitles, ultra-natural voiceover, voice cloning, the whole pipeline in one go. Chinese-English translation is especially stable, Whisper recognition accuracy is ridiculously high, LLM translates segment by segment without losing context, CosyV…
Introducing KrillinAI, a free locally-run video translation tool that supports precise subtitles, natural voiceover, and voice cloning. It integrates Whisper, LLM, and CosyVoice, and supports Win/Mac and yt-dlp.
@GitTrend0x: Holy cow, guys! Run voice cloning and cinematic video dubbing locally, supporting 646 languages, fully offline, no API key, no internet needed. ElevenLabs is crushed! https://github.com/debpalash/OmniVoice-Studio… This open-source marvel is insane...
OmniVoice Studio is an open-source desktop app that enables local voice cloning and cinematic video dubbing across 646 languages, fully offline with no API keys, positioning itself as a privacy-focused alternative to ElevenLabs.