2026
RLVER: Reinforcement Learning with Verifiable Emotion Rewards for Empathetic Agents
ICLR 2026poster
Large language models (LLMs) excel at logical and algorithmic reasoning, yet their emotional intelligence (EQ) still lags far behind their cognitive prowess. While reinforcement learning from verifiable rewards (RLVR) has advanced in other domains, its application to dialogue—especially for emotion…