Llama 4 Scout 17B-16E Instruct

Meta

Catalogue only Open weights

Sovereignty and hosting

Catalogue only

This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.

Summary

Llama 4 Scout is a natively multimodal mixture-of-experts model with 17B active parameters, 16 experts and 109B total parameters. Meta advertises a 10 million token context window for it, the largest of any Llama release, and states it fits on a single H100 with Int4 quantisation. It is released under the Llama 4 Community License, which keeps the 700M monthly-active-user threshold.

Specifications

Provider
Meta (US)
Context window
10.5M
Parameters
108.64B
Modalities
text, vision
License
other · licence
Open weights
yes
Release date
05 Apr 2025
Tasks
vision, summarization, rag, multilingual-support

Languages

Strong: en, de, fr, es, it, pt, hi

Adequate: ar, id, th, tl, vi

Languages

Strong

  • en
  • de
  • fr
  • es
  • it
  • pt
  • hi

Adequate

  • ar
  • id
  • th
  • tl
  • vi

Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.

When to use it

Strengths

  • 10M-token context window, the longest published in the Llama line
  • multimodal input with image grounding
  • fits a single H100 at Int4 quantisation

Limits

  • Llama 4 Community License, not an OSI licence
  • 109B total parameters make full-precision serving expensive
  • long-context quality at the 10M limit is not independently verified