← Search

Xiao Deng

1 accepted papers

2025

Can You Really Trust Code Copilot? Evaluating Large Language Models from a Code Security Perspective

ACL 2025long

Code security and usability are both essential for various coding assistant applications driven by large language models (LLMs). Current code security benchmarks focus solely on single evaluation task and paradigm, such as code completion and generation, lacking comprehensive assessment across dimen…