Next generation medical image interpretation with MedGemma 1.5 and medical speech to text with MedASR
Google Research
Google released MedGemma 1.5 4B, an updated medical AI model that now handles high-dimensional medical imaging like CT and MRI scans alongside text and 2D images, and launched MedASR, a speech-to-text model fine-tuned for medical dictation. MedGemma 1.5 improved accuracy by 3% on CT classification and 14% on MRI classification over its predecessor, while MedASR achieved 58% fewer errors than Whisper on chest X-ray dictations. Google is hosting a $100,000 MedGemma Impact Challenge hackathon and making both models freely available on Hugging Face and Google Cloud for developers to build healthcare applications.
Why it matters
Generative AI