1
Hidden in Plain Sight: Benchmarking Agent Safety Against Decomposition Attacks with DECOMPBENCH
针对AI代理安全的分解攻击,提出DECOMPBENCH基准测试框架,揭示隐藏风险。
arXiv:2606.13994v1 Announce Type: cross Abstract: LLM-based Agents are becoming increasingly capable and widely deployed, creating growing incentives …