Tag
Vois 2.0 is a new AI voice generation tool launched as an alternative to ElevenLabs, offering unlimited generation capabilities.
Fish Audio launches S2.1 Pro, a production voice model with 90ms latency, support for 83 languages, voice cloning from short samples, and multi-speaker dialogue, available via API with a free tier for development.
Fish Audio announced its S2.1 Pro model, which can clone a voice from just 5 seconds of audio, alongside a $52M seed funding round.
Fish Audio has raised $52 million in seed funding to develop AI voice models for creators and enterprises. The startup, which generates $21M in ARR and has 8 million users, offers open-source and paid voice generation models, including its latest S2.1 Pro API.
The article discusses the significance of voice technology in artificial intelligence.
A comparison of three voice AI agents — Retell, Vapi, and Plura AI — evaluating their performance for production use cases.
A detailed, unsponsored review of ElevenLabs rating it 8.1/10, highlighting its emotional range and low latency as strengths, but cautioning about high costs and the 'regeneration tax' for casual users and high-volume publishers.
OpenAI's new GPT voice model enables highly realistic, real-time voice conversations with low latency and emotional expression, marking a significant leap in AI voice interaction.
OpenAI has finally delivered the original promise of AVM, enabling advanced voice interactions.
GPT-Live, as an AI speaking teacher, can correct grammar and unidiomatic expressions in real time, potentially having a significant impact on English speaking instruction.
Netflix uses AI from ElevenLabs to recreate Gene Wilder's voice for its upcoming reality show Wonka's The Golden Ticket, with consent from the actor's family.
A leaked version of ChatGPT featuring Bidi-1 voice mode sounds eerily realistic, surpassing previous leaks.
Kokoro-82M is a highly natural text-to-speech model with 82 million parameters and over 11 million downloads, representing a significant advancement in AI voice generation.
Introducing VoxCPM2, a completely free for commercial use, open-source multilingual voice synthesis model supporting voice design, cloning, and 48kHz high-quality output, ranked #1 on GitHub trending.
ElevenLabs signs a deal with Stan Lee Universe to create an AI clone of Stan Lee's voice and likeness for digital cameos, audiobooks, and a book club series, sparking ethical debates about consent and exploitation.
GitHub open-source project VoxCPM2 achieves AI voice cloning without reference audio, generating target voice precisely with just one sentence, has gained 20K stars.
Google announces Gmail Live, an AI voice mode for searching and interacting with your inbox via Gemini, along with similar voice-driven features for Docs and Keep, rolling out summer 2026 to AI Pro/Ultra subscribers.
DramaBox by Resemble AI converts scene descriptions into AI-generated vocal performances.
Elon Musk announces that Grok Voice has reached the number one ranking.
Vapi_ai announces a $50M Series B funding round led by Peak XV Partners, totaling $72M raised, with a focus on engineering for AI voice calls.