← Search

Ofek Glick

1 accepted papers

2025

MATCH: Task-Driven Code Evaluation through Contrastive Learning

EMNLP 2025

AI-based code generation is increasingly prevalent, with GitHub Copilot estimated to generate 46% of the code on GitHub. Accurately evaluating how well generated code aligns with developer intent remains a critical challenge. Traditional evaluation methods, such as unit tests, are often unscalable a