chatbot-safety

Tag

Cards List
#chatbot-safety

GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety

arXiv cs.AI · 2026-05-22 Cached

GrandGuard introduces a comprehensive taxonomy, benchmark, and safeguards for elderly-specific risks in LLM chatbot interactions, finding that leading LLMs mishandle over 50% of such risks and proposing two safeguards achieving up to 96.2% detection accuracy.

0 favorites 0 likes
#chatbot-safety

VERA-MH: Validation of Ethical and Responsible AI in Mental Health

arXiv cs.AI · 2026-05-14 Cached

VERA-MH is a clinically-validated evaluation framework to assess the safety of chatbots in mental health support, focusing on suicidal ideation risks. It uses role-played conversations and an LLM-as-a-Judge with a clinical rubric to evaluate responses.

0 favorites 0 likes
← Back to home

Submit Feedback