Tag
A tutorial demonstrating a low-latency voice agent setup with Pipecat PhoneLLM Alpha 1 on Modal, using Deepgram transcription and Cartesia voice, with code and video guides.
Deepgram released Flux TTS, a streaming conversation-native text-to-speech model that retains tone, pacing, and context across turns, handles interruptions, and runs with latency as low as 80ms to make voice AI feel more natural.