voice-cloning

Tag

Cards List
#voice-cloning

Qwen3.8-LiveTranslate: Names the speaker. Carries the meaning (3 minute read)

TLDR AI ↗ · 2026-09-21

Qwen3.8-LiveTranslate is an AI model that reduces translation lag and improves quality using an interleaved audio-text architecture, featuring real-time speaker separation and voice cloning in 60 languages.

0 favorites 0 likes
#voice-cloning

@sentient_agency: holy shit... GitHub is insane right now developers are casually open-sourcing AI agents, browser automation, voice clon…

X AI KOLs Timeline ↗ · 2026-09-18 Cached

The tweet showcases ten trending open-source GitHub repositories focused on AI agents, browser automation, voice cloning, and local model execution, indicating a growing infrastructure for AI agent development.

0 favorites 0 likes
#voice-cloning

@nikola_mr64990: this is f**king insane. a solo dev just open sourced a 100% FREE ElevenLabs replacement that runs entirely on your own …

X AI KOLs Timeline ↗ · 2026-09-13 Cached

A solo developer has open-sourced a free, local AI tool that serves as an ElevenLabs replacement, enabling voice cloning, video dubbing in 646 languages, and more without any data leaving the user's machine.

0 favorites 0 likes
#voice-cloning

@Ryrenz: So powerful—OpenVoice just needs a short reference audio clip to clone that voice timbre, and it can even switch langua…

X AI KOLs Timeline ↗ · 2026-09-12 Cached

OpenVoice is an open-source AI tool for voice cloning that uses a short reference audio clip to replicate timbre, enabling multilingual and emotional adjustments, and has gained significant traction on GitHub.

0 favorites 0 likes
#voice-cloning

Steven Johnson (Google Labs) just described exactly how to clone a writer's voice without asking

Reddit r/artificial ↗ · 2026-09-12

Steven Johnson of Google Labs explains how AI can clone a writer's voice without permission, highlighting challenges for copyright protection and the freelance economy.

0 favorites 0 likes
#voice-cloning

LoudKit: local TTS with voice cloning, 10 languages, and SDKs for Python, Swift, Go, Rust and TypeScript

Reddit r/LocalLLaMA ↗ · 2026-09-10

LoudKit is an open-source local TTS tool with voice cloning, supporting 10 languages, and available as SDKs for Python, Swift, Go, Rust, and TypeScript, designed to run efficiently on edge devices.

0 favorites 0 likes
#voice-cloning

@TeksEdge: Tencent open-sourced a small 1.5B AI model that can replace a whole stack of separate audio tools. Beats Qwen3-TTS! Thi…

X AI KOLs Timeline ↗ · 2026-09-10 Cached

Tencent has open-sourced a 1.5B parameter AI model called AuK that can replace multiple audio tools, handling tasks like TTS, voice cloning, and denoising via natural language instructions.

0 favorites 0 likes
#voice-cloning

Neural Vocals: Voice Cloning & Timbre Transfer · Trust Node Logic

Reddit r/artificial ↗ · 2026-09-10 Cached

The article discusses three converging toolchains in neural voice synthesis—vocal identity, signal manipulation, and performance realism—focusing on voice cloning and timbre transfer to transform vocal performances while preserving original phrasing and dynamics.

0 favorites 0 likes
#voice-cloning

Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech

arXiv cs.CL ↗ · 2026-09-04 Cached

This paper presents a method to build a compact fixed-voice Thai TTS system using synthetic speech from a larger model, evaluating its performance and introducing an 82M-parameter model for on-device deployment.

0 favorites 0 likes
#voice-cloning

Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech

Hugging Face Daily Papers ↗ · 2026-09-03 Cached

This paper presents Wayu-Paxa-TTS-Edge, an 82M-parameter Thai TTS model trained on synthetic speech from a voice-cloning teacher, achieving high accuracy and prosody for on-device use without reference audio.

0 favorites 0 likes
#voice-cloning

@ErickSky: The project I'm bringing you tonight is an ElevenLabs... but local and in 646 languages. [VoiceStudio] It's the open-so…

X AI KOLs Timeline ↗ · 2026-09-01 Cached

VoiceStudio is an open-source local voice studio tool offering voice cloning, dubbing, transcription, and more in 646 languages, providing an alternative to ElevenLabs.

0 favorites 0 likes
#voice-cloning

@Huahuazo: After trying this voice cloning tool, I was stunned. I casually took an 8-second recording of myself speaking, threw it in, and the output voice—the tone, pauses, that cheeky tone—was exactly like me arguing. The best part is it runs completely locally, no audio upload needed, privacy and security are fully maximized. Honestly, my first reaction wasn't 'impressive,' but rather 'chills down my spine': in the future, just by a voice recording, you really can't take things too seriously…

X AI KOLs Timeline ↗ · 2026-08-30 Cached

Introducing the local voice cloning tool VoiceStudio, which only requires 8 seconds of recording to highly replicate the voice, runs completely locally to protect privacy, and supports multiple languages and platforms.

0 favorites 0 likes
#voice-cloning

@leeoxiang: A very lightweight TTS implementation, attempted to replicate Audio8's training from scratch using 2000 hours of data. Trained on H200 for less than 10 hours, and the SIM metric already reached 0.72.

X AI KOLs Timeline ↗ · 2026-08-29 Cached

A lightweight TTS implementation, replicating Audio8's training with 2000 hours of data, achieving a SIM metric of 0.72 in under 10 hours of training on H200.

0 favorites 0 likes
#voice-cloning

TontaubeV1 - Open TTS model release for local long-form generation

Reddit r/LocalLLaMA ↗ · 2026-08-28

TontaubeV1 is an open-weight text-to-speech model released for local long-form generation, supporting English and German with zero-shot voice cloning and low-latency inference on GPUs.

0 favorites 0 likes
#voice-cloning

Fast On-Device Voice Cloning (9 minute read)

TLDR AI ↗ · 2026-08-28 Cached

Sopro V2 is an open-source, fast, on-device text-to-speech model with voice cloning, specifically targeting European Portuguese and other languages for low-latency and private communication.

0 favorites 0 likes
#voice-cloning

Best AI Voice Cloning in 2026: How to Clone Your Voice With AI

Reddit r/LocalLLaMA ↗ · 2026-08-24 Cached

The article reviews and ranks AI voice cloning models from 2026, including CosyVoice 3 and VibeVoice, based on their accuracy in replicating the author's voice through fine-tuning and zero-shot methods.

0 favorites 0 likes
#voice-cloning

Building and Evaluating a Synthetic Bengali Speech Resource for Telecom Customer Care

arXiv cs.CL ↗ · 2026-08-24 Cached

This paper introduces a synthetic Bengali speech dataset of 10,000 audio-text pairs for telecom customer care scenarios, generated using OmniVoice voice-cloning, and evaluates it with an ASR model, achieving low word error rates.

0 favorites 0 likes
#voice-cloning

@IndieDevHailey: VoiceStudio: local version of ElevenLabs, voice cloning, sound design, video dubbing, dictation, audiobooks, all included, supporting 646 languages. 16 TTS engines switchable at will, completely local, no account needed, no API, no subscription, all data on your machine. Just renamed from OmniVoice, already has 1…

X AI KOLs Timeline ↗ · 2026-08-22 Cached

VoiceStudio is a locally running AI voice tool, offering features such as voice cloning, sound design, video dubbing, etc., supporting 646 languages, with no need for network connection or subscription, and all data processed locally.

0 favorites 0 likes
#voice-cloning

FireRedAudio & FireRedTTS3 by FireRedTeam - Huggingface

Reddit r/LocalLLaMA ↗ · 2026-08-21

FireRedAudio and FireRedTTS3 are AI models for unified audio understanding and generation, offering ASR, TTS, voice cloning, editing, and multilingual support with competitive benchmarks.

0 favorites 0 likes
#voice-cloning

Audio8/Audio8-TTS-Preview-0.1b

Hugging Face Models Trending ↗ · 2026-08-19 Cached

Audio8 TTS Preview 0.1b is a compact zero-shot text-to-speech model with approximately 170M parameters for the main model, supporting voice cloning and multiple languages.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback