1
AIPatient Arena: EHR-grounded evaluation of large language models in end-to-end clinical consultation workflows
首个基于电子健康记录评估大模型在临床咨询全流程表现的研究,为医学AI落地提供量化标尺。
arXiv:2606.17474v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly considered for use in clinical consultation tasks, yet…