Introducing Gemma 3n: The developer guide
Google DeepMind ● Covered by 4 sources
Google released Gemma 3n, a mobile-focused AI model family designed to run multimodal applications directly on edge devices with minimal memory requirements. The E4B version achieves an LMArena score over 1300, making it the first model under 10 billion parameters to reach this benchmark, while requiring as little as 2GB to 3GB of memory for operation. The model's MatFormer architecture enables developers to extract custom-sized variants between two pre-built sizes, with the E2B sub-model offering up to 2x faster inference than the larger E4B version.
Why it matters
Gemma 3n is designed for the developer community that helped shape Gemma.