The Fragility of Strategic Thinking in Large Language Models
不只是问答,这项研究揭示了LLM在战略博弈中的推理短板,值得关注。
arXiv:2510.10813v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly applied to domains that require reasoning about othe…
不只是问答,这项研究揭示了LLM在战略博弈中的推理短板,值得关注。
arXiv:2510.10813v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly applied to domains that require reasoning about othe…
LLM策略短板明显,记忆增强代理却能让推理能力飙升,值得关注。
arXiv:2608.12626v1 Announce Type: cross Abstract: Strategic reasoning in Large Language Models (LLMs) within long-horizon environments is often limite…
AI推理能力提升引发战略推理风险,新论文提出分类评估框架应对欺骗与评估游戏
arXiv:2604.22119v2 Announce Type: replace Abstract: As reasoning capacity and deployment scope grow in tandem, large language models (LLMs) gain the c…
用扑克对弈平台多维度剖析LLM战略推理与记忆能力,揭示大模型在复杂决策场景的真实水平。
arXiv:2606.13815v1 Announce Type: new Abstract: Strategic reasoning under uncertainty underpins consequential decisions in negotiation, finance, and p…
论文提出MINDGAMES,一个实时竞技场,用于评估多智能体LLM的社交与战略推理能力,突破静态基准局限。
arXiv:2605.29512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as interactive agents, yet their capacity for s…
LLM在多智能体游戏中因对手策略非平稳性而表现不佳,Strat-Reasoner通过强化战略推理显著提升决策能力。
arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent …