We built NeuTTS-2E, an open-source on-device TTS model with 7 controllable emotions
Summary
NeuTTS-2E is an open-source on-device TTS model that supports seven controllable emotions.
Similar Articles
EmoRES-TTS: Residual-Enhanced Vector Steering for Emotional Speech Generation
The paper introduces EmoRES-TTS, a training-free method for emotional speech generation that enhances controllability by decomposing emotion vectors into shared and residual components, achieving superior performance over existing methods on benchmarks.
Introducing Eleven v4, our most emotive model (9 minute read)
ElevenLabs has launched Eleven v4, their most emotive text-to-speech AI model, which excels in generating natural, context-aware speech with emotional depth and a low-latency variant for real-time applications.
Adding emotion control tags to Qwen3-TTS
The article describes fine-tuning Qwen3-TTS to add emotion control tags, overcoming training challenges like codec prefix inconsistencies and generation concurrency issues, and discovering that emotion can be manipulated via affine transformations in speaker embeddings.
@multimodalart: they extracted only the audio bit of LTX-2.3, fine-tuned for TTS task and achieved SOTA TTS emotional control??? try it…
A fine-tuned version of the LTX-2.3 model's audio component achieves state-of-the-art emotional control in text-to-speech, now available as a Hugging Face Space called DramaBox by ResembleAI.
What is the best tts to create audio books? One that support emotions?
A user is asking for recommendations on the best text-to-speech technology that supports emotional expression for creating audiobooks, citing dissatisfaction with current options like Google TTS.