1
I-CALM: Incentivizing Confidence-Aware Abstention for LLM Selective Answering
让大模型学会“不知道就承认”,用激励机制打造可靠的选择性回答,直击幻觉痛点。
arXiv:2604.03904v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often produce confident but incorrect answers, in part because …