lucataco/xtts-v2
Summary
XTTS-v2 is an open foundation speech model by Coqui that supports multiple languages for text-to-speech synthesis, with recent additions like Hindi support.
View Cached Full Text
Cached at: 09/14/26, 02:20 AM
Similar Articles
TontaubeV1 - Open TTS model release for local long-form generation
TontaubeV1 is an open-weight text-to-speech model released for local long-form generation, supporting English and German with zero-shot voice cloning and low-latency inference on GPUs.
Zyphra/ZONOS2
ZONOS2 is a new text-to-speech model from Zyphra trained on over 6 million hours of multilingual speech, offering high-quality voice cloning and low latency using a mixture-of-experts architecture. It supports 30+ languages and includes a high-performance inference server.
kyutai-labs/pocket-tts
Kyutai releases Pocket TTS, a lightweight text-to-speech model that runs efficiently on CPUs with 100M parameters, low latency, and voice cloning, supporting multiple languages.
CVSS-X: A Multilingual Speech-to-Speech Translation Corpus for 28 Languages
CVSS-X is a large-scale synthetic speech-to-speech translation corpus extending CVSS to enable translation from English into 28 languages, with over 16,000 hours of parallel speech pairs for bidirectional research.
Higgs Audio v3 TTS 4B. Built for voice chat. Support 100 languages and inline control.
Higgs Audio v3 is a 4B parameter TTS model designed for voice chat applications, supporting 100 languages with inline control capabilities.