AffectOmni trains multimodal AI to explain emotions using people-centered visual evidence
A new preprint describes AffectOmni, a reinforcement-learning framework designed to make multimodal AI models base affective judgments on people-centered cues such as facial expressions, body language and temporal order. The authors report improvements over open-source 7B-scale baselines on three benchmarks…