SyncAI.news, a Varaisys broadcasting
Gemma 4: Byte for byte, the most capable open models
GD

Google DeepMind

· 1 min read

AI LabsGoogle DeepMind

Gemma 4: Byte for byte, the most capable open models

Apr 02, 2026

|

Clement Farabet

VP of Research, Google DeepMind

Olivier Lacombe

Director Product Management, Google Deepmind

Your browser does not support the audio element.

Listen to article

[[duration]] minutes

This content is generated by Google AI. Generative AI is experimental

Today, we are introducing Gemma 4 — our most intelligent open models to date. Purpose-built for advanced reasoning and agentic workflows, Gemma 4 delivers an unprecedented level of intelligence-per-parameter. This breakthrough builds on incredible community momentum: since the launch of our first generation, developers have downloaded Gemma over 400 million times, building a vibrant Gemmaverse of more than 100,000 variants. We listened closely to what innovators need next to push the boundaries of AI, and Gemma 4 is our answer: breakthrough capabilities made widely accessible under an Apache 2.0 license.

Open model performance vs size on Arena.ai’s chat arena as of 4/1.

Built from the same world-class research and technology as Gemini 3, Gemma 4 is the most capable model family you can run on your hardware. They complement our Gemini models, giving developers the industry's most powerful combination of both open and proprietary tools.

Industry-leading capabilities and mobile-first AI

We are releasing Gemma 4 in four versatile sizes: Effective 2B (E2B), Effective 4B (E4B), 26B Mixture of Experts (MoE) and 31B Dense. The entire family moves beyond simple chat to handle complex logic and agentic workflows. Our larger models deliver state-of-the-art performance for their sizes, with the 31B model currently ranking as the #3 open model in the world on the industry-standard Arena AI text leaderboard, and the 26B model securing the #6 spot. There, Gemma 4 outcompetes models 20x its size. For developers, this new level of intelligence-per-parameter means achieving frontier-level capabilities with significantly less hardware overhead.

Original source

This story was published by Google DeepMind. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on deepmind.google

Similar News