conversational-ai

Tag

Cards List
#conversational-ai

@btaylor: I'm proud to announce our partnership with SoftBank, the leading telecommunications and IT operator in Japan, with deca…

X AI KOLs Following · 2026-07-14 Cached

Sierra announces a strategic partnership with SoftBank to deliver its conversational AI platform to Japanese enterprises, with SoftBank serving as exclusive sales partner and adopting the platform across its own brands.

0 favorites 0 likes
#conversational-ai

ChatGPT-Live vs Pi vs Lucy OS1 vs Gemini-Live: best AI assistant to talk with?

Reddit r/artificial · 2026-07-11

A comparison of AI voice assistants ChatGPT-Live, Pi, Lucy OS1, and Gemini-Live focusing on which feels most natural to talk with, concluding that conversational quality is becoming the key differentiator as intelligence improves.

0 favorites 0 likes
#conversational-ai

The complexities of patient-centred conversational artificial intelligence

arXiv cs.AI · 2026-07-10 Cached

This paper analyzes 2,053 real patient-chatbot conversations to show that communication styles vary widely and can significantly alter triage outcomes, finding that patient simulators that model emotional state and conversational strategy produce conversations nearly indistinguishable from real ones in a Turing test.

0 favorites 0 likes
#conversational-ai

OpenAI releases new voice models for more natural live conversations

TechCrunch AI · 2026-07-08 Cached

OpenAI released new full-duplex voice models GPT-Live-1 and GPT-Live-1 mini for more natural live conversations, allowing simultaneous speaking and listening, with improvements in turn-taking and context handling, and replacing Advanced Voice Mode in ChatGPT.

0 favorites 0 likes
#conversational-ai

ChatGPT’s upgraded voice mode is better at shutting up

The Verge · 2026-07-08 Cached

OpenAI releases GPT-Live-1, an upgraded voice mode for ChatGPT that interrupts less, allows real-time translation, and can generate visual context for topics like weather and sports.

0 favorites 0 likes
#conversational-ai

From Passive Retrieval to Active Memory Navigation: Learning to Use Memory as a Structured Action Space

arXiv cs.AI · 2026-07-08 Cached

NapMem is a framework that treats long-term user memory as a structured action space rather than passive retrieval, using a multi-granularity memory pyramid and reinforcement learning to train agents to navigate memory. Experiments show competitive performance on memory-intensive tasks.

0 favorites 0 likes
#conversational-ai

I spent a while trying to get an LLM to make a podcast that's actually listenable. The hard part wasn't the model.

Reddit r/artificial · 2026-07-07

A developer shares techniques for making LLM-generated podcasts sound natural, including using constraints to force disagreement and pre-editing content before generation.

0 favorites 0 likes
#conversational-ai

Don't Wait to Reply: Towards Responsive yet Thoughtful Dialogue through Proactive Thinking

arXiv cs.CL · 2026-07-07 Cached

Proposes a Proactive Thinking framework that allows LLMs to pre-compute response elements during conversational pauses, improving interaction efficiency without sacrificing quality. Introduces a training-free baseline that speculatively anticipates future states, evaluated on time-aware benchmarks.

0 favorites 0 likes
#conversational-ai

Your own custom conversational AI podcast with two hosts that you can interrupt to ask questions in real time

Reddit r/ArtificialInteligence · 2026-07-04

A new tool lets you create a custom conversational AI podcast with two hosts that you can interrupt to ask questions in real time.

0 favorites 0 likes
#conversational-ai

@omarsar0: So basically, prompt engineering isn't dead. Jokes aside, great read! You leave a lot on the table if you don't prompt …

X AI KOLs Following · 2026-07-03 Cached

A tweet highlights that prompt engineering remains relevant, recommending a read on how to enrich AI agent interactions through techniques like brainstorming and planning.

0 favorites 0 likes
#conversational-ai

@_philschmid: "Make it day time." The lighting shift, the shadows move, the sky changes. Gemini Omni Flash can edit your videos throu…

X AI KOLs Following · 2026-07-02 Cached

Google's Gemini Omni Flash model can edit videos through conversational prompts, using the Interactions API to generate new clips based on user descriptions.

0 favorites 0 likes
#conversational-ai

Learning User-Aware Recall: Personalized Retrieval in Long-Term Conversational Memory

arXiv cs.AI · 2026-07-02 Cached

This paper introduces Profile-guided Personalized Retrieval Optimization (PPRO), a framework that enhances long-term conversational agents by incorporating user profiles into memory retrieval and optimizing retrieval via reinforcement learning, achieving consistent improvements over existing methods.

0 favorites 0 likes
#conversational-ai

TRACE: State-Aware Query Processing over Temporal Evidence Graphs for Conversational Data

arXiv cs.CL · 2026-07-02 Cached

This paper presents TRACE, a query processing framework that models conversational data as temporal evidence graphs to enable state-aware reasoning over evolving user states, improving temporal and multi-hop reasoning for long-conversation QA.

0 favorites 0 likes
#conversational-ai

Reference-Based Prosody and Rhythm Evaluation for Spoken Dialogue Systems

arXiv cs.CL · 2026-07-01 Cached

This paper proposes a reference-based evaluation protocol for assessing prosody and rhythm in speech-to-speech AI systems, using matched human conversation data to provide interpretable behavioral plausibility checks.

0 favorites 0 likes
#conversational-ai

IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations

arXiv cs.AI · 2026-06-30 Cached

IMCBench is a new benchmark for evaluating multimodal LLMs on image-grounded medical conversations, pairing clinical images with synthetic patient profiles. Evaluations across safety, accuracy, and uncertainty show that even strong models like Claude Opus 4.6 have safety issues, highlighting the need for multi-dimensional evaluation.

0 favorites 0 likes
#conversational-ai

From Lexicon to AI: A Structured-Data Pipeline for Specialized Conversational Systems in Low-Resource Languages

arXiv cs.CL · 2026-06-26 Cached

Presents a systematic methodology for converting Hindi WordNet into 1.25 million instruction-response pairs to fine-tune a 12B-parameter language model using LoRA, demonstrating improved pedagogical effectiveness for specialized conversational systems in low-resource languages.

0 favorites 0 likes
#conversational-ai

Reducing Conversational Escalation in Large Language Model Dialogue with Nonviolent Communication Constraints

arXiv cs.CL · 2026-06-26 Cached

This paper explores using Nonviolent Communication (NVC) principles as lightweight prompt constraints to reduce conversational escalation in LLMs during conflict-prone interactions. Experiments across multiple instruction-tuned models show that NVC-constrained prompting consistently de-escalates dialogue and stabilizes interactions with highly resistant users.

0 favorites 0 likes
#conversational-ai

Dziri Voicebot: An End-to-End Low-Resource Speech-to-Speech Conversational System for Algerian Dialect

arXiv cs.CL · 2026-06-25 Cached

This paper presents a modular end-to-end speech-to-speech conversational system for the low-resource Algerian Dialect, integrating ASR, NLU, RAG, and TTS with dedicated datasets and fine-tuned models.

0 favorites 0 likes
#conversational-ai

@bnicholehopkins: Overwhelmed by the support on our Series A announcement this morning, including the incredible piece by Chris Metinko f…

X AI KOLs Following · 2026-06-24 Cached

Coval announces a Series A funding round to build infrastructure for testing and deploying conversational AI agents in enterprises, inspired by the rigor of autonomous vehicle testing.

0 favorites 0 likes
#conversational-ai

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

Hugging Face Daily Papers · 2026-06-23 Cached

This paper introduces a conversational voice agent system that uses a lightweight on-device 'Talker' model to start responding immediately, then incorporates knowledge from a frontier LLM 'Reasoner' as it becomes available, achieving 7-19x faster time-to-first-response while approaching frontier-level performance on a laptop.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback