Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning
难度感知熵正则化,让大模型推理时自动压缩简单问题算力,集中资源攻克难题,效率与精度兼得。
arXiv:2602.22642v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has substantially empowered Large Language Models (LLMs) to tackle complex …