voice-assistants

Tag

Cards List
#voice-assistants

ChatGPT-Live vs Pi vs Lucy OS1 vs Gemini-Live: best AI assistant to talk with?

Reddit r/artificial · 2026-07-11

A comparison of AI voice assistants ChatGPT-Live, Pi, Lucy OS1, and Gemini-Live focusing on which feels most natural to talk with, concluding that conversational quality is becoming the key differentiator as intelligence improves.

0 favorites 0 likes
#voice-assistants

Full duplex vs half duplex - the spectrum of AI voice models [D]

Reddit r/MachineLearning · 2026-06-01

An analysis of half-duplex vs full-duplex architecture in AI voice models, discussing key features like overlap, backchannels, and barge-in that make voice agents sound robotic.

0 favorites 0 likes
#voice-assistants

Gemini Spark next week and Siri 2.0 in two weeks are the last serious shots at making AI agents a consumer product

Reddit r/ArtificialInteligence · 2026-05-23

Google's Gemini Spark and Apple's Gemini-powered Siri 2.0 are launching in the next two weeks, representing major attempts to bring AI agents to mainstream consumers with billions of potential users.

0 favorites 0 likes
#voice-assistants

Three bots in a trenchcoat is not omnichannel

Reddit r/AI_Agents · 2026-05-11

Elba showcases its unified AI architecture that bridges voice and text channels in real-time, contrasting its single-agent system with competitors' fragmented 'three bots in a trenchcoat' approaches.

0 favorites 0 likes
#voice-assistants

MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models

arXiv cs.CL · 2026-04-20 Cached

MoshiRAG combines a compact full-duplex speech language model with asynchronous retrieval-augmented generation to improve factuality while maintaining real-time interactivity. The approach leverages natural temporal gaps in conversation to retrieve external knowledge without disrupting the natural flow of dialogue.

0 favorites 0 likes
← Back to home

Submit Feedback