1
Calibration Without Comprehension: Diagnosing the Limits of Fine-Tuning LLMs for Vulnerability Detection in Systems Software
微调大模型检测系统漏洞,看似精准实则可能只是校准错觉,揭示能力边界的关键研究。
arXiv:2606.20502v1 Announce Type: cross Abstract: Whether LLMs scoring well on vulnerability benchmarks genuinely reason about security or merely patt…