Gemma 4 E4B IT
Sovereignty and hosting
This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.
Summary
Gemma 4 E4B IT is the smallest Gemma 4 model with native audio, handling text, images and audio input with a 128K-token context window. Google states that the "E" in E2B and E4B stands for effective parameters and that these sizes use Per-Layer Embeddings, so the stored model is about 8B parameters while the effective count is smaller. The model card reports 69.4% on MMLU Pro and 58.6% on GPQA Diamond, and it is Apache-2.0 licensed.
Specifications
- Provider
- Google DeepMind (US)
- Context window
- 131k
- Parameters
- 8B
- Modalities
- text, vision, audio
- License
- apache-2.0 · licence
- Open weights
- yes
- Release date
- 02 Mar 2026
- Tasks
- cheap, vision, multilingual-support, summarization
Languages
Strong: en, de, fr, es, it, pt
Adequate: nl, pl, sv, cs, ro, el, uk, ru, ar, hi, zh, ja, ko, tr
Languages
Strong
- en
- de
- fr
- es
- it
- pt
Adequate
- nl
- pl
- sv
- cs
- ro
- el
- uk
- ru
- ar
- hi
- zh
- ja
- ko
- tr
Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.
When to use it
Strengths
- native audio input in the small-model class
- Apache-2.0 licence
- runs on modest hardware
- text, image and audio in one checkpoint
Limits
- reasoning scores well below the larger Gemma 4 sizes
- 128K-token window, half of the medium models
- no evaluation has been run by LLM EU