TLDRocket
Sign in

Speech Recognition

14 summarised stories about Speech Recognition, each linking back to the original source. Browse all topics →

+ Follow this topic

Tuesday, 13 January 2026

Next generation medical image interpretation with MedGemma 1.5 and medical speech to text with MedASR

Google Research 7 months ago 39

Google released MedGemma 1.5 4B, an updated medical AI model that now handles high-dimensional medical imaging like CT and MRI scans alongside text and 2D images, and launched MedASR, a speech-to-text model fine-tuned for medical dictation. MedGemma 1.5 improved accuracy by 3% on CT classification and 14% on MRI classification over its predecessor, while MedASR achieved 58% fewer errors than Whisper on chest X-ray dictations. Google is hosting a $100,000 MedGemma Impact Challenge hackathon and making both models freely available on Hugging Face and Google Cloud for developers to build healthcare applications.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.