Tag
A developer built OpenClaw Voice, a dedicated voice interface using Home Assistant Voice PE and OpenAI Realtime, enabling sub-second smart home control, memory-based queries, phone calling, and long-running tasks with voice feedback. The entire stack is open-sourced.
Whispera 是一个基于 VoxCPM 的 Windows 本地实时语音助手,集成了 SenseVoice ASR、llama-server 本地 LLM 推理、VoxCPM 流式 TTS 和 Mem0 长期记忆,完全离线运行。
Amazon has announced an agentic version of Alexa with long-term memory and integration with over 1,000 apps, enhancing its capabilities as a personal AI assistant.
Explores the performance of running a voice assistant with Qwen3-ASR and Kokoro-TTS ONNX models on CPU, measuring response times without a GPU.
Athena is a fully offline, privacy-first voice assistant running entirely on local hardware, combining multiple AI models for speech recognition, language understanding, and emotion-aware speech synthesis, with long-term memory and interruptibility.
nxt is a tool that lets you talk to your to-do list to determine what to do next.
Hands-on review of the Google Home Speaker, praising its sound quality and design but noting finicky touch controls and hidden light ring.
Proposes BindingSubspace (BSU), a representation-level framework that isolates and attenuates intent-conditioned directions in end-to-end spoken language understanding models to prevent capability persistence, where suppressing an intent still allows slot generation under forced prefixes. The method reduces forced-prefix recoverability while preserving retained performance on SLU benchmarks.
Google Home is getting an update that improves facial recognition by using clothing and body size to identify people, and can now recognize specific sounds like dog barking or alarms in video event descriptions.
Amazon is testing a Hindi-language version of its generative AI assistant Alexa+ in India, inviting users to join a beta program to refine the experience before a wider launch.
Wired's hands-on with Apple's Siri AI on iOS 27 shows a more conversational, personalized, and helpful assistant, powered by Google's Gemini and Apple Intelligence, with plans for public release later this year.
Google announces the $99.99 Google Home Speaker, its first standalone smart speaker in years, powered by Gemini AI to enable natural language and multistep commands, with a $10/month premium subscription for advanced features.
Google's Gemini TTS now supports streaming audio generation, allowing developers to build voice applications that start speaking instantly without waiting for full audio output.
Google has finally launched its new Gemini-powered Home Speaker, six years after its last smart speaker, offering improved natural language understanding, multi-command handling, and a redesigned form factor. Preorders start June 17, with official sales on June 25 for $100.
VoiceOS is a voice assistant that acts like JARVIS for your computer, enabling hands-free control.
The Verge's first 24 hours with Siri AI on macOS 27 Golden Gate beta, finding it still limited but showing promise for basic tasks like app launching while lacking deeper app integration.
Apple has decided not to launch Siri in the European Union after its request for an exemption from certain EU digital regulations was denied, impacting the rollout of its voice assistant in the region.
Apple announced 'Siri AI' at WWDC, a more conversational voice assistant with deeper integration across apps and personal context, rolling out this fall alongside Google-powered AI model updates.
Wired's updated 2026 guide to the best smart speakers across Amazon Alexa, Google Assistant, and Apple Siri ecosystems, highlighting top picks like the Amazon Echo Dot Max, Google Nest Hub Max, Google Nest Audio, and Apple HomePod Mini.
Nvidia CEO Jensen Huang announced at Computex 2026 that the company is planning N2X and N3X chips, envisioning a future of Star Trek-like AI computers that users can talk to and remotely control.