1
Thinking at the Right Size: Amortized Distillation Across Post-Trained LLMs
不同尺寸模型间摊销蒸馏,让推理成本与思考深度灵活匹配,EMNLP 2026 前沿成果
arXiv:2608.22854v1 Announce Type: new Abstract: Practical deployment of large language models (LLMs) requires families of post-trained variants---inst…