← Search

Aadim Nepal

1 accepted papers

2025

Warm Up Before You Train: Unlocking General Reasoning in Resource-Constrained Settings

EMNLP 2025

Designing effective reasoning-capable LLMs typically requires training using Reinforcement Learning with Verifiable Rewards (RLVR) or distillation with carefully curated Long Chain of Thoughts (CoT), both of which depend heavily on extensive training data. This creates a major challenge when the amo