Annual Meeting of the Association for Computational Linguistics
A.S.E: A Repository-Level Benchmark for Evaluating Security in AI-Generated Code
The paper introduces A.S.E (AI Code Generation Security Evaluation), a repository-level benchmark for assessing the security of AI-generated code. It is built from 120 instances derived from 40 real-world GitHub repositories with documented CVEs, expanded via semantic and structural mutations. The benchmark covers four vulnerability types (SQL injection,…
Keke Lian, Bin Wang, Lei Zhang, Libo Chen, et al.- Published
- Aug 2025
- Upvotes
- 350
- Citations
- 15