1
MIRAGE: Auditing Anti-Muslim Bias in Frontier LLMs Across Reasoning, Agentic, and Time-Coupled Conditions
前沿LLM反穆斯林偏见评估基准MIRAGE,覆盖推理、智能体与时间耦合场景,揭示模型偏见新维度。
arXiv:2606.16562v1 Announce Type: new Abstract: Five years after the discovery of persistent anti-Muslim bias in large language models, most evaluatio…