Position: Evaluations of AI Moral Reasoning Still Miss Half of the Picture
AI道德评估只测了一半?新研究揭示当前评测盲区,搞AI对齐的都该看看
arXiv:2608.14566v1 Announce Type: new Abstract: Recent work on evaluating the moral competence of large language models (LLMs) has focused primarily o…