Google documents Gemma 4 with MoE, thinking, and 256K context
Google's launch post presents Gemma 4 as a five-size open-model family with E2B, E4B, 12B, 26B A4B, and 31B variants; the 26B A4B model is a mixture-of-experts option with roughly 4B active parameters per token, image-and-text input, text output, thinking, function calling, and up to 256K context.