multi-turn-conversation

Tag

Cards List
#multi-turn-conversation

TAF-MED: Multi-Turn Safety Refusal Collapse in LLMs Under Declared Self-Treatment Intent

arXiv cs.CL · 4d ago Cached

Introduces TAF-MED, a physician-reviewed benchmark of 500 multi-turn medical safety scenarios, showing that LLMs often collapse from safe initial refusals to unsafe responses when users declare self-treatment intent. Evaluation of eight LLMs across 4,000 conversations finds 71.6% contained unsafe responses and first-turn safety is an insufficient proxy for conversational safety persistence.

0 favorites 0 likes
#multi-turn-conversation

LLM-OSDA: An Optimal-Stopping Dynamic Auction for Native Advertising in Multi-Turn LLM Conversations

arXiv cs.CL · 2026-08-04 Cached

This paper introduces LLM-OSDA, a dynamic cost-per-click auction for native advertising in multi-turn LLM conversations, integrating Bellman optimal stopping, winner allocation, and envelope pricing. Experiments show an 11% net revenue improvement over fixed-timing baselines while maintaining user retention.

0 favorites 0 likes
#multi-turn-conversation

Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation

arXiv cs.CL · 2026-05-26 Cached

This paper proposes PUMA, a framework for LLM personalization in multi-turn conversations that models latent user states and uses the Free Energy Principle to select dialogue actions, improving long-horizon outcomes on healthcare counseling benchmarks.

0 favorites 0 likes
#multi-turn-conversation

@HowToAI_: Microsoft Research + Salesforce has published a paper that should scare every single AI builder right now. It’s called …

X AI KOLs Timeline · 2026-05-09

A new paper by Microsoft Research and Salesforce reveals that LLM performance drops significantly in multi-turn conversations due to a 'Lost in Conversation' phenomenon, challenging the reliability of current single-turn benchmarks.

0 favorites 0 likes
← Back to home

Submit Feedback