Open (Apache 2.0) family of multimodal AI models from Google DeepMind (E2B/E4B/26B A4B/31B). Supports text, image, audio, and video. Native function calling.
Context window
256K
tokens
Parameters
25.2B
parameters
Access:APIDownloadHostedDeployment:๐ป Localโ Cloud๐ฑ On-device
Overview
Access & deployment
APIDownloadHosted
LocalCloudOn-device
Weights: Open source
Key parameters
๐ Context: 256K
๐งฉ Parameters: 25.2B
โ Toolsย ยทย โ Fine-tuning
๐ฅ Input: text, image, audio, video
Platforms
Technical specification
Context window
256K
tokens
Parameters
25.2B
parameters
License
Apache 2.0
Hardware requirements
E2B/E4B: mobile and edge devices (phones, tablets, IoT); 26B A4B (MoE): consumer GPU or workstation; 31B: workstation-class GPU. The E2B/E4B variants are designed to run on devices without cloud connectivity.
Features:โ Tool useโ Fine-tuning
Modalities
โฌ Input
textimageaudiovideo
โฌ Output
textcodestructured_data
Capabilities and applications
Native model capabilities
Coding
Generating, analysing and modifying code in many programming languages. Covers writing functions, debugging, refactoring, code review, and creating tests. Measured by benchmarks such as HumanEval and SWE-bench.
Category: coding
Multilingual
Competence in many natural languages (from a few to over a hundred): understanding, generation, translation, and code-switching within a single conversation. Frontier models support a wide range of languages with comparable quality.
Category: language
Multi-step reasoning
Carrying out multi-step chains of reasoning across long, complex tasks.
Category: reasoning
Long context
Support for large context windows โ tens to hundreds of thousands (or millions) of input tokens. Enables analysis of entire codebases, long documents, and many parallel conversations without losing earlier information. GPT-5.1 supports 400,000 tokens.
Category: language
Structured output
Producing data in structured formats such as JSON.
Category: structured_generation
Image understanding
Analysing and interpreting the content of images.
Category: vision
Audio understanding
Category: audio
Multimodal understanding
Category: multimodal
Reasoning
The model's ability to reason logically and solve complex problems.
Category: reasoning
Video Understanding
Category: video
Function Calling
Category: planning
Interleaved Multimodal Input
Category: reasoning
Benchmark results
5 benchmarks
MMLU Pro
accuracy ยท instruction-tuned (Gemma 4 31B IT)
85.2%
๐
31 Mar 2026๐ Gemma 4 model card | Google AI for Developers
Source: ai.google.dev/gemma/docs/core/model_card_4. Score for Gemma 4 31B (IT). MMLU Pro is harder than standard MMLU.
GPQA
accuracy ยท Instruction-tuned variant (Gemma 4 31B IT).
84.3%
๐
31 Mar 2026๐ Gemma 4 model card | Google AI for Developers
Source: ai.google.dev/gemma/docs/core/model_card_4. Result for Gemma 4 31B (IT). Diamond subset of GPQA.
LiveCodeBench v6
accuracy ยท instruction-tuned (Gemma 4 31B IT)
80.0%
๐
31 Mar 2026๐ Gemma 4 model card | Google AI for Developers
Source: ai.google.dev/gemma/docs/core/model_card_4. Result for Gemma 4 31B (IT). LiveCodeBench v6 coding benchmark.
AIME 2026 (no tools)
accuracy ยท Instruction-tuned variant (Gemma 4 31B IT), no tools enabled.
89.2%
๐
31 Mar 2026๐ Gemma 4 model card | Google AI for Developers
Source: ai.google.dev/gemma/docs/core/model_card_4. Score for Gemma 4 31B (IT). American Invitational Mathematics Examination 2026.
MMMU Pro (Vision)
accuracy ยท Instruction-tuned variant (Gemma 4 31B IT), vision-capable model.
76.9%
๐
31 Mar 2026๐ Gemma 4 model card | Google AI for Developers
Source: ai.google.dev/gemma/docs/core/model_card_4. Score for Gemma 4 31B (IT). MMMU Pro is an extended, more challenging version of MMMU.
Pricing
Technical architecture
Core Architecture
Model Form
Training Techniques
Deployment and security
โ Available on platforms
๐ Security / Enterprise
โ Verified enterprise information
Gemma 4 model card includes safety evaluation results. As an open-source model, deployment responsibility lies with the user. Documentation on responsible AI use is available.
Gemma 4 is an open-source model (Apache 2.0). Production deployments require independent risk assessment. Google publishes a model card with safety evaluation results.
Updated: 5 Apr 2026โ Security documentation
Sources and related pages
8 sources
DocsGemma 4 โ przeglฤ
d modelu | Google AI for DevelopersPaperGemma 4 model card | Google AI for DevelopersBlogGemma 4: Byte for byte, the most capable open models | Google BlogDocsApache License 2.0 | GemmaRepogoogle/gemma-4-31B-it | Hugging FaceRepogoogle/gemma-4-26B-A4B-it | Hugging FaceWebGemma | Google DeepMindDocsGemma releases | Google AI for Developers
