Tag
Introduces TAF-MED, a physician-reviewed benchmark of 500 multi-turn medical safety scenarios, showing that LLMs often collapse from safe initial refusals to unsafe responses when users declare self-treatment intent. Evaluation of eight LLMs across 4,000 conversations finds 71.6% contained unsafe responses and first-turn safety is an insufficient proxy for conversational safety persistence.