2025
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
NeurIPS 2025poster
Recent reports claim that large language models (LLMs) now outperform elite humans in competitive programming. Drawing on knowledge from a group of medalists in international algorithmic contests, we revisit this claim, examining how LLMs differ from human experts and where limitations still remain.…