1
Vibe Coding on Trial: Operating Characteristics of Unanimous LLM Juries
当单个大模型评代码不够可信,让多个LLM组成「陪审团」一致表决会怎样?实验数据揭示真相
arXiv:2602.18492v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are now good enough at coding that developers can describe inte…