Mistral Large 3 (675B)
Mistral AI · Providers EU
Sovereignty and hosting
This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.
Summary
Mistral Large 3 is the largest published Mistral model: a granular mixture-of-experts model with 675B total parameters and 41B active parameters, released under Apache-2.0. It is multimodal and supports a 256k-token context window. The vendor states it was trained from scratch on 3000 H200 GPUs.
Specifications
- Provider
- Mistral AI (FR)
- Context window
- 262k
- Parameters
- 675B
- Modalities
- text, vision
- License
- apache-2.0 · licence
- Open weights
- yes
- Release date
- 28 Nov 2025
- Tasks
- reasoning, coding, multilingual-support, vision, summarization
Languages
Strong: en, fr, es, de, it, pt
Adequate: nl, zh, ja, ko, ar
Languages
Strong
- en
- fr
- es
- de
- it
- pt
Adequate
- nl
- zh
- ja
- ko
- ar
Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.
When to use it
Strengths
- frontier-scale open weights under a permissive licence
- 41B active parameters keep decoding cheaper than the total size suggests
- 256k-token context window
- vendor-published NVFP4 checkpoint for single-node deployment
Limits
- 675B parameters require multi-GPU nodes even quantised to NVFP4
- no independent evaluation has been run by LLM EU
- vendor benchmark comparisons are self-reported