In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores
打破AI公平性评测的“应试教育”,用真实行为模式替代标准化测试分数。
arXiv:2605.12530v2 Announce Type: replace-cross Abstract: LLM fairness should be evaluated through in-situ behavioral pattern rather than standardized…