2026
FeatureBench: Benchmarking Agentic Coding for Complex Feature Development
ICLR 2026poster
Agents powered by large language models (LLMs) are increasingly adopted in the software industry, contributing code as collaborators or even autonomous developers. As their presence grows, it becomes important to assess the current boundaries of their coding abilities. Existing agentic coding benchm…