Gemma 4 E4B IT

Google DeepMind

Catalogue only Open weights

Sovereignty and hosting

Catalogue only

This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.

Summary

Gemma 4 E4B IT is the smallest Gemma 4 model with native audio, handling text, images and audio input with a 128K-token context window. Google states that the "E" in E2B and E4B stands for effective parameters and that these sizes use Per-Layer Embeddings, so the stored model is about 8B parameters while the effective count is smaller. The model card reports 69.4% on MMLU Pro and 58.6% on GPQA Diamond, and it is Apache-2.0 licensed.

Specifications

Provider
Google DeepMind (US)
Context window
131k
Parameters
8B
Modalities
text, vision, audio
License
apache-2.0 · licence
Open weights
yes
Release date
02 Mar 2026
Tasks
cheap, vision, multilingual-support, summarization

Languages

Strong: en, de, fr, es, it, pt

Adequate: nl, pl, sv, cs, ro, el, uk, ru, ar, hi, zh, ja, ko, tr

Languages

Strong

  • en
  • de
  • fr
  • es
  • it
  • pt

Adequate

  • nl
  • pl
  • sv
  • cs
  • ro
  • el
  • uk
  • ru
  • ar
  • hi
  • zh
  • ja
  • ko
  • tr

Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.

When to use it

Strengths

  • native audio input in the small-model class
  • Apache-2.0 licence
  • runs on modest hardware
  • text, image and audio in one checkpoint

Limits

  • reasoning scores well below the larger Gemma 4 sizes
  • 128K-token window, half of the medium models
  • no evaluation has been run by LLM EU