2025
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
EMNLP 2025
Large language models excel at complex tasks by breaking down problems into structured reasoning steps. However, reasoning traces often extend beyond reaching a correct answer, causing wasted computation, reduced readability, and hallucinations. To address this, we introduce a novel hyperparameter-f