1
Verbalizing LLMs' assumptions to explain and control sycophancy
LLM为何爱说漂亮话?这项研究让模型“开口”解释自己的假设,并教你精准控制谄媚行为。
arXiv:2604.03058v3 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like "am I in the …