@yhslgg: Bro, sharing another open-source video translation tool—pyVideoTrans, with 17,700 stars on GitHub, a must-have for video repurposing and localization! In a nutshell: drop a video in, and it automatically runs through the entire pipeline of speech recognition → subtitle translation → AI dubbing → video synthesis, outputting a complete video in another language. Core...
Summary
pyVideoTrans is an open-source video translation tool that supports automatic speech recognition, subtitle translation, AI dubbing, and video synthesis. It integrates multiple ASR, translation, and TTS engines, making it suitable for cross-language video production and localization.
View Cached Full Text
Cached at: 06/03/26, 07:54 PM
Guys, sharing another open-source video translation tool — pyVideoTrans, 17,700 stars on GitHub. A must-have for video repurposing and localization!
In a nutshell: Drop a video in, it automatically runs through the full pipeline — speech recognition → subtitle translation → AI voiceover → video synthesis — and out comes the complete video in another language.
Core advantages, one by one:
(1) Speaker diarization — Can handle multi-speaker videos, distinguishing different voices so subtitles don’t get mixed up. Works for interviews, variety shows, courses.
(2) Voice cloning — Integrates F5-TTS, CosyVoice, GPT-SoVITS. You can clone a specific voice for dubbing instead of using generic synthetic audio, giving a more natural, human-like result.
(3) Pause for manual review at each stage — Check after recognition, edit after translation. No need to fully trust automation; quality control stays in your hands.
(4) Comprehensive tech stack, easy to swap — ASR supports Faster-Whisper, Alibaba Qwen, Azure, Google; Translation supports DeepSeek, Claude, Gemini, ChatGPT, Ollama local models; TTS offers Edge-TTS (free), OpenAI, Azure, ChatTTS, ChatterBox. Pick your preferred tool for each step.
(5) GPU acceleration — Supports NVIDIA CUDA and AMD GPUs, maxing out processing speed so you don’t waste time waiting.
(6) CLI headless mode — Command-line support, ideal for batch processing on servers, enabling automated pipelines.
(7) Windows portable exe version — No need to set up a Python environment. Just download and run, very user-friendly for casual users.
Who it’s for: Cross-language content repurposers, adding Chinese subtitles to foreign content, taking Chinese videos global by translating into English/Japanese/Korean, or translating English courses into Chinese for personal use — efficiency doubles instantly.
Similar Articles
@Huahuazo: The most annoying part of reposting overseas videos? Manually cutting subtitles, eye-straining translation alignment, and guessing audio sync—this workflow is so inefficient it makes you want to smash your keyboard. VideoLingo connects the entire process into an automated pipeline. It has 7k+ stars on GitHub and an MIT license, so it's reliable to use. …
VideoLingo is an open-source video translation, localization, and dubbing tool that uses WhisperX and AI to deliver Netflix-level subtitles and multilingual dubbing. It supports downloading via yt-dlp and multiple TTS options, helping video reposters automate the entire workflow.
@yhslgg: Old Yang shares another gem open-source tool—KrillinAI, 10,000 stars on GitHub, a must-see for multilingual audio/video content! In a nutshell: from video download to subtitle translation, AI dubbing, video compositing, the entire pipeline is covered, and it can even auto-generate platform covers, supporting Bilibili, Douyin, Xiaohongshu, YouTube…
KrillinAI is an open-source tool that integrates the entire workflow of video downloading, subtitle translation, AI dubbing, and video compositing. It supports context-aware translation, voice cloning, auto layout, and cover generation, and is compatible with multiple AI models, suitable for multilingual audio/video content creation and distribution.
@XAMTO_AI: If you don't bookmark this open-source tool now, you'll regret it later — automatic video dubbing and translation, supports 33 languages at once, and can even answer questions about video content. Found a gem on GitHub called Violin, fully open-source, what it does is a bit unbelievable: you drop a video in, it automatically recognizes speech, …
Violin is an open-source automatic video dubbing and translation tool that supports 33 languages, integrates models like Whisper and DeepSeek, and provides one-click speech recognition, translation, dubbing synthesis, and in-video Q&A functionality.
@CycleDecoded: Unbelievable, guys, someone actually made this—it's directly taking away the livelihood of video re-uploaders. Just found an open-source magical tool on GitHub that quietly makes money: Y2A-Auto. This thing is an emotionless automated pipeline that can fully automatically download YouTube videos, process them in one go, and post to …
Y2A-Auto is an open-source, fully automatic video repurposing tool that downloads videos from YouTube, uses AI for translation, subtitle generation, and content moderation, and automatically uploads to Bilibili and AcFun for unattended operation.
@berryxia: Guys, this is awesome! Install it right away! Kevin Lin, postdoc at Oxford, former Meta and Microsoft researcher, just released Violin, an open-source video translation Skill. Video is already the absolute dominant content form on the internet. Yet most high-quality lectures, speeches, and podcasts are locked by a single language…
Violin is an open-source video translation tool that integrates speech recognition, large language model translation, and text-to-speech. It supports over 30 languages and offers three usage modes: CLI, web app, and Claude Code.