Llama 4 Scout 17B-16E Instruct
Sovereignty and hosting
This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.
Summary
Llama 4 Scout is a natively multimodal mixture-of-experts model with 17B active parameters, 16 experts and 109B total parameters. Meta advertises a 10 million token context window for it, the largest of any Llama release, and states it fits on a single H100 with Int4 quantisation. It is released under the Llama 4 Community License, which keeps the 700M monthly-active-user threshold.
Specifications
- Provider
- Meta (US)
- Context window
- 10.5M
- Parameters
- 108.64B
- Modalities
- text, vision
- License
- other · licence
- Open weights
- yes
- Release date
- 05 Apr 2025
- Tasks
- vision, summarization, rag, multilingual-support
Languages
Strong: en, de, fr, es, it, pt, hi
Adequate: ar, id, th, tl, vi
Languages
Strong
- en
- de
- fr
- es
- it
- pt
- hi
Adequate
- ar
- id
- th
- tl
- vi
Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.
When to use it
Strengths
- 10M-token context window, the longest published in the Llama line
- multimodal input with image grounding
- fits a single H100 at Int4 quantisation
Limits
- Llama 4 Community License, not an OSI licence
- 109B total parameters make full-precision serving expensive
- long-context quality at the 10M limit is not independently verified