论文

通用AI与临床AI评测基准受限

Limited benchmarks constrain the conclusions of a general-purpose versus clinical AI comparison

精选理由

Nature Medicine最新研究揭示,现有评测基准不足以全面比较通用AI和临床AI模型的表现差异。

Nature Medicine发表研究指出,当前有限的评测基准导致通用AI模型与临床AI模型的比较结论存在局限性。研究发表于2026年9月3日,DOI编号为10.1038/s41591-026-04638-6。该研究强调了建立更全面、专业评测基准的必要性。

图片来源 · Nature Medicine
原文 · Nature Medicine

Limited benchmarks constrain the conclusions of a general-purpose versus clinical AI comparison

Nature Medicine, Published online: 03 September 2026; doi:10.1038/s41591-026-04638-6 Limited benchmarks constrain the conclusions of a general-purpose versus clinical AI comparison