IssueTrojanBench: Benchmarking AI Coding Agents Against Malicious Issue Requests
首个针对AI编程代理的恶意issue请求基准测试,揭示AI agent安全漏洞与鲁棒性挑战。
arXiv:2607.20759v1 Announce Type: cross Abstract: AI coding agents powered by LLMs are increasingly integrated into real-world software development, w…