2023
Aligning Offline Metrics and Human Judgments of Value for Code Generation Models
ACL 2023findings
Large language models have demonstrated great potential to assist programmers in generating code. For such human-AI pair programming scenarios, we empirically demonstrate that while generated code are most often evaluated in terms of their functional correctness (i.e., whether generations pass avail…