rl-verification

Tag

Cards List
#rl-verification

AffectOmni: RL-Verifiable People-Centric Grounded Affective Reasoning for Social and Art-Related Scenes

arXiv cs.AI · 2026-08-28 Cached

AffectOmni is a GRPO-trained framework for verifiable affective reasoning in multimodal large language models, introducing People Focus and Temporal Order rewards to enhance people-centric evidence selection and temporally structured reasoning, with experiments showing improvements over 7B scale baselines.

0 favorites 0 likes
← Back to home

Submit Feedback