Tag
EQ-Bench 4 is a 16-turn chat benchmark that assesses AI chatbots' emotional intelligence by simulating diverse user personas with competing traits, measuring perception, adaptability, and trust repair in roleplay situations.
This paper presents PTEI, a framework that integrates personality traits (MBTI and OCEAN) into LLMs to enhance emotional intelligence, using contrastive learning and personality-aware prompts. Experiments show significant improvements in emotional understanding, especially when combined with Chain-of-Thought reasoning.
This paper evaluates four leading real-time voice AI systems (GPT Realtime 2, Gemini 3.1 Flash Live, Qwen3.5 Omni Plus, Omni Flash) and finds they consistently act on words rather than vocal tone, ignoring distress, fear, or sarcasm even when they can perceive them—termed the 'emotional intelligence gap' of voice AI.
SpeechEQ introduces a benchmark and dataset for evaluating emotional intelligence in speech-language models, covering 15 EQ subscales across 2,265 dialogues. Experiments reveal current models struggle with paralinguistic cues, exhibiting text-reliant shortcuts and other limitations.
A tweet pondering whether the first generation raised with AI will become superhumans or emotionally disconnected, sparking debate on AI's societal impact.