1
Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception
大模型越自信越容易说谎?这篇论文揭示置信度如何放大LLM的欺骗风险,对AI安全研究有重要启示。
arXiv:2607.20444v1 Announce Type: cross Abstract: Large language models (LLMs) can produce deceptive responses: outputs that mislead users in service …