1
SAGE: A Search-AuGmented Evaluation of Large Language Models on Free-Form QA
搜索增强评测框架SAGE,严谨检验大模型在开放问答中的真实表现,值得AI研究者一读。
arXiv:2504.07385v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) become increasingly used for question-answering (QA), relyin…