1
Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol
面对LLM代理频繁修订判断标准,这项研究揭示失败模式并提出追踪锚定协议,值得AI研究者细读。
arXiv:2608.20729v1 Announce Type: new Abstract: Language-model agents can improve after failure or carry text across episodes without revising what co…