Tail-Aware Information-Theoretic Bounds for LLM Alignment under Heavy-Tailed Rewards
从理论层面重新审视LLM对齐的奖励分布假设,揭示重尾效应对对齐边界的影响,值得关注。
arXiv:2604.10727v2 Announce Type: replace-cross Abstract: Classical information-theoretic learning bounds typically rely on KL mutual information and …