1
Uncertainty-Aware Abstention in Large Language Models with Provable Alignment Guarantees
让大模型在不确定时主动拒答,并给出可证明的安全对齐保证,缓解幻觉与越狱风险。
arXiv:2607.04430v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in question answering (QA) systems, yet they ma…