2026
R4: Nested Reasoning-Retrieval for Reward Modeling in Role-Playing Agents
ICLR 2026poster
Role-playing dialogue presents unique challenges for large language models (LLMs): beyond producing coherent text, models must sustain character persona, integrate contextual knowledge, and convey emotional nuance. Despite strong reasoning abilities, current LLMs often generate dialogue that is lite…