emotion-recognition

Tag

Cards List
#emotion-recognition

Exposing Weaknesses in Emotion Recognition in Conversations

arXiv cs.AI · 16h ago Cached

This paper investigates weaknesses in emotion recognition in conversations (ERC) by analyzing LLM performance in zero-shot settings, revealing systematic failures due to annotation ambiguity, and proposes an LLM-as-Judge framework for more robust evaluation.

0 favorites 0 likes
#emotion-recognition

Chiaroscuro for Emotions: A Contrastive Emotion Benchmark Grounded in Appraisal Theory

arXiv cs.CL · 6d ago Cached

The paper presents Chiaro, a new benchmark dataset for contrastive emotion recognition where two individuals experience opposing emotions from a shared event, grounded in appraisal theory. It evaluates seven LLMs and four emotion classifiers, revealing that current models fall short of human performance.

0 favorites 0 likes
#emotion-recognition

VoiceLongMemEval: Do Assistants Remember How You Sounded?

arXiv cs.AI · 2026-09-02 Cached

The paper introduces VoiceLongMemEval (VLME), a benchmark that evaluates AI assistants' ability to remember and reason over paralinguistic metadata like emotion and prosody from voice in long-term conversations, revealing an 'affect gap' in current models.

0 favorites 0 likes
#emotion-recognition

VocalAffectBench: Evaluating Vocal Emotion Recognition in AI Audio Models

arXiv cs.CL · 2026-09-01 Cached

VocalAffectBench is introduced as a public benchmark for evaluating vocal emotion recognition in AI audio models, demonstrating that current baselines have limited accuracy, particularly for non-neutral emotions.

0 favorites 0 likes
#emotion-recognition

AffectOmni: RL-Verifiable People-Centric Grounded Affective Reasoning for Social and Art-Related Scenes

arXiv cs.AI · 2026-08-28 Cached

AffectOmni is a GRPO-trained framework for verifiable affective reasoning in multimodal large language models, introducing People Focus and Temporal Order rewards to enhance people-centric evidence selection and temporally structured reasoning, with experiments showing improvements over 7B scale baselines.

0 favorites 0 likes
#emotion-recognition

@nrol_ling: Seven languages, 4,421 utterances: is emotion represented the same way in all of them? Mostly yes. The geometry lines u…

X AI KOLs Timeline · 2026-08-24 Cached

The research finds that emotion representation is mostly universal across seven languages in speech models, with a measurable 'accent' that mirrors human cross-cultural studies and affects cross-lingual transfer.

0 favorites 0 likes
#emotion-recognition

Emotion Across Speech and Faces: Shared Affective Mechanisms in Multimodal Foundation Models

arXiv cs.CL · 2026-08-19 Cached

This research paper explores emotion-sensitive neurons in multimodal foundation models, revealing shared affective mechanisms between speech and facial emotion recognition through causal interventions and cross-modal analysis.

0 favorites 0 likes
#emotion-recognition

Rationale-Guided Learning for Multimodal Emotion Recognition

arXiv cs.AI · 2026-08-12 Cached

Introduces Rationale-Guided Learning (RGL), a framework that reframes multimodal emotion recognition in conversation as a cognitively-inspired reasoning task using dual-process theory and MLLM-generated rationales, achieving state-of-the-art results on IEMOCAP and MELD.

0 favorites 0 likes
#emotion-recognition

CONFER: Conflict-Aware Evidence Negotiation for Regime-Calibrated Weak Supervision in Multimodal Emotion Recognition

arXiv cs.LG · 2026-08-11 Cached

This paper proposes CONFER, a graph-based conflict-aware evidence negotiation framework for weakly supervised multimodal emotion recognition, addressing self-report unreliability and cross-modal conflict. It achieves competitive accuracy on AMIGOS, MAHNOB-HCI, and DEAP benchmarks.

0 favorites 0 likes
#emotion-recognition

Separating Decision-Rule Misalignment from Readout-Coverage Limitations in Speech Language Models

arXiv cs.CL · 2026-08-10 Cached

This preprint introduces a generation-aligned diagnostic ladder that separates decision-rule misalignment from readout-coverage limitations in speech language models, showing that state decoding far exceeds generated accuracy in emotion recognition tasks.

0 favorites 0 likes
#emotion-recognition

C$^2$MOE: Consistency and Complementarity-guided Mixture of Experts for Incomplete Multimodal Emotion Learning

arXiv cs.AI · 2026-08-06 Cached

The paper proposes C²MOE, a Consistency and Complementarity-guided Mixture of Experts framework for incomplete multimodal emotion recognition in conversations, using information-theoretic decomposition to improve robustness when modalities are missing.

0 favorites 0 likes
#emotion-recognition

Evaluation Protocols and Cross-Subject Generalization in EEG Emotion Recognition

arXiv cs.LG · 2026-07-31 Cached

This paper examines how evaluation protocols affect reported accuracy in EEG emotion recognition, using a DGCNN on SEED and SEED-IV datasets. It demonstrates that subject-dependent, subject-disjoint, and cross-session evaluations answer different questions, and that checkpoint selection and test-set reuse can inflate accuracy.

0 favorites 0 likes
#emotion-recognition

AtmosERC: Modeling Dialogue-Level Affective Atmosphere for Emotion Recognition in Conversation

arXiv cs.CL · 2026-07-30 Cached

This paper introduces AtmosERC, a model that models dialogue-level affective atmosphere to enhance emotion recognition in conversations.

0 favorites 0 likes
#emotion-recognition

Dynamic Commonsense Coordination for Empathetic Response Generation

arXiv cs.CL · 2026-07-27 Cached

Proposes DCC, a dynamic commonsense coordination framework for empathetic response generation that integrates residual-based interaction, association-guided filtering, and iterative decoding, achieving improved emotion classification and response diversity over baselines.

0 favorites 0 likes
#emotion-recognition

SCoPE: Shift-Aware Speaker-Conditioned Priors for Emotion Recognition in Conversations

arXiv cs.CL · 2026-07-24 Cached

Introduces SCoPE, a lightweight module for emotion recognition in conversations that models speaker-specific emotional priors and uses emotion shift prediction to dynamically fuse prior and multimodal evidence, achieving state-of-the-art on IEMOCAP.

0 favorites 0 likes
#emotion-recognition

Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?

arXiv cs.AI · 2026-07-15 Cached

The paper proposes Light-MER, a lightweight multimodal emotion recognition framework that uses knowledge distillation from an 8B teacher model to a sub-1B student, achieving state-of-the-art performance with significantly higher inference efficiency, challenging the necessity of models larger than 1B parameters.

0 favorites 0 likes
#emotion-recognition

Graph-Regularized Deep Learning for EEG-Based Emotion Recognition with Psychologically-Grounded Label Structure

arXiv cs.LG · 2026-07-10 Cached

The paper introduces a graph-regularized deep learning framework for EEG-based emotion recognition that incorporates psychologically-grounded emotion topology into the training objective, achieving up to +5.42% accuracy and 39% reduction in psychologically implausible misclassifications on SEED datasets.

0 favorites 0 likes
#emotion-recognition

SHAP-Weighted Cross-Modal Expert Fusion for Emotion and Sentiment Recognition: Evidence and Limits

arXiv cs.AI · 2026-07-10 Cached

This paper proposes SHAP-weighted cross-modal expert fusion (XGAF) for emotion and sentiment recognition, demonstrating that sum-abs SHAP aggregation achieves early-fusion-level performance on MELD and CMU-MOSEI datasets.

0 favorites 0 likes
#emotion-recognition

@OrukLabs: None of these models was ever told what emotion is. They were trained to transcribe words, or to fill in masked audio. …

X AI KOLs Following · 2026-07-04 Cached

A study from OrukLabs shows that speech models trained solely on transcription or masked audio tasks spontaneously learn to represent emotions in their deeper layers, as revealed by mapping with real voice clips.

0 favorites 0 likes
#emotion-recognition

PRISM: Prioritized Channel Importance with Semi-supervised Domain Adaptation for Cross-Subject EEG Emotion Recognition

arXiv cs.LG · 2026-07-02 Cached

PRISM is a novel framework for cross-subject EEG emotion recognition that combines prioritized channel importance weighting via a lightweight expert ensemble with semi-supervised domain adaptation using confidence-filtered pseudo-labels, achieving state-of-the-art results on DEAP, DREAMER, and SEED datasets.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback