2026
Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning
ICML 2026poster
We introduce **Native Parallel Reasoner (NPR)**, a teacher-free framework that enables Large Language Models (LLMs) to self-evolve genuine parallel reasoning capabilities. NPR transforms the model from sequential emulation to native parallel cognition through three key innovations: 1) a **self-disti…