@omarsar0: Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers car…
Summary
Miso One is an open-source 8B parameter text-to-speech model with real emotional range and 110ms latency, designed for voiceover work.
View Cached Full Text
Cached at: 06/03/26, 05:54 PM
Another banger open-source release.
Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation, and excitement instead of sounding flat.
It’s purpose-built for voiceover work like shorts, podcasts, and educational content, and it runs at 110ms latency, which is faster than human reaction time.
The best part is that the weights are fully open source, so you can clone the repo, self-host, fine-tune, and keep your data private.
Worth checking out if you’re building voice into your tools and products: http://github.com/MisoLabsAI/MisoTTS…
Aoden Teo (@AodenTeoMT): Today, we’re excited to introduce Miso One, the most emotive voice model in the world.
Miso One is an 8-billion-parameter text-to-speech model for highly expressive speech generation. It emotes like a human and responds faster than a human, with just 110 milliseconds of
Similar Articles
MisoLabs/MisoTTS
Miso Labs releases Miso TTS 8B, a text-to-speech model based on the Sesame CSM architecture with a Llama 3.2-style backbone, designed for high-quality conversational speech generation and voice continuation.
OpenMOSS-Team/MOSS-TTS-v1.5 · Hugging Face
MOSS-TTS v1.5 is an updated open-source text-to-speech model with improved multilingual synthesis (supporting 31 languages), more stable zero-shot voice cloning, and explicit inline pause control.
OpenMOSS-Team/MOSS-TTS-Nano-100M
MOSS-TTS-Nano is an open-source multilingual speech generation model with only 0.1B parameters, designed for real-time TTS that runs directly on CPU without GPU. Released by OpenMOSS team and MOSI.AI, it enables simple local deployment for web serving and product integration.
@Open_MOSS: MOSS-Transcribe-Diarize is among the top trending models on @huggingface . A few weeks after release, the model has bee…
MOSS-Transcribe-Diarize, an open-source ASR model with multi-speaker diarization and hotword biasing, trends on Hugging Face after release.
@Gorden_Sun: ZONOS2: Open-source MoE TTS model. 8B total parameters, 0.9B activated parameters. Supports multilingual, voice cloning, Chinese, and Chinese results are good. Model:
Zyphra released ZONOS2, an open-source MoE text-to-speech model trained on over 6 million hours of multilingual speech, supporting voice cloning and high-quality synthesis across many languages.