1
Learning When to Sample: Confidence-Aware Selective Sampling for Efficient Chain-of-Thought Reasoning
通过自适应置信度评估选择性采样推理路径,在保持CoT性能的同时大幅降低计算成本,为高效大模型推理提供新思路。
arXiv:2603.08999v3 Announce Type: replace Abstract: Large language models (LLMs) can achieve strong reasoning performance through chain-of-thought (Co…