Gemma
Google’s first lightweight open models, released in 2B and 7B sizes.
FAMILY / GOOGLE DEEPMIND
Google’s open-model family for local, edge, and customizable deployment.
01 / LINEAGE MAP
02 / GENERATIONS
Google’s first lightweight open models, released in 2B and 7B sizes.
Raised reasoning quality in 9B and 27B sizes designed for accessible accelerators.
Added image input, a 128K context window, and support for more than 140 languages.
Used MatFormer and Per-Layer Embeddings to bring multimodal understanding to mobile devices.
The high-quality dense Gemma 4 flagship for local reasoning, coding, and agent workflows.
The Gemma 4 MoE activating only 3.8B parameters per step for local speed and intelligence density.
A unified encoder-free multimodal model for laptops, filling the gap between E4B and 26B.
An experimental Gemma 4 text-diffusion model exploring lower latency through parallel block generation.