Can You Really Trust Code Copilot? Evaluating Large Language Models from a Code Security Perspective
Code security and usability are both essential for various coding assistant applications driven by large language models (LLMs). Current code security benchmarks focus solely on single evaluation task and paradigm, such as code completion and generation, lacking comprehensive assessment across dimen…