1
Implicit Reasoning Steering via Concept Chaining
利用概念链引导大模型进行隐式推理,为提升推理可控性提供新思路
arXiv:2607.14242v1 Announce Type: new Abstract: Large language models often appear to reason reliably, yet on many questions repeated sampling yields …
利用概念链引导大模型进行隐式推理,为提升推理可控性提供新思路
arXiv:2607.14242v1 Announce Type: new Abstract: Large language models often appear to reason reliably, yet on many questions repeated sampling yields …
强化学习推理新思路:不直接给答案,引导模型在rollout中自省纠错,提升推理深度。
arXiv:2510.09388v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has become a key driver for enhancing the long chain-of-thought (CoT) …
新方法通过智能体引导链式思维,让大模型推理更高效、更可控。
arXiv:2606.03965v1 Announce Type: cross Abstract: Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but…