1
Is Lying an Emergent Behaviour in LLMs? Evidence from Gaslighting AI agents in a Sustainability Game
大模型会在游戏中自发撒谎并“煤气灯”他人?这项研究用可持续性博弈给出了惊人证据,关乎AI安全与对齐。
arXiv:2606.28456v1 Announce Type: cross Abstract: LLMs agents are increasingly used in multi-agent settings, yet their behaviour in sustainability gam…