模型精选

MedGemma医疗视觉语言模型发布

An open vision-language model for diverse medical applications

精选理由

谷歌发布了MedGemma医疗模型,能处理医学影像和文本,比同类模型表现更好。

MedGemma是基于Gemma 3的医疗视觉语言基础模型集合。该模型在多个医学影像领域展现出先进的医学理解和推理能力。其性能超越了同等规模的生成式模型,同时保持了Gemma基础模型的通用能力。

图片来源 · Nature Medicine
原文 · Nature Medicine

An open vision-language model for diverse medical applications

Nature Medicine, Published online: 06 October 2026; doi:10.1038/s41591-026-04626-w MedGemma, a collection of medical vision-language foundation models based on Gemma 3, demonstrates advanced medical understanding and reasoning across images and text and multiple medical imaging domains, exceeding the performance of similarly sized generative models while maintaining the general capabilities of the Gemma base models.