European languages
Modellqualität auf übersetzten europäischen Benchmarks. Die EU21-Zeilen stammen aus der OpenGPT-X-Evaluation, die mit Teuken 7B Instruct v0.6 veröffentlicht wurde und ARC, HellaSwag, TruthfulQA und MMLU in die 21 abgedeckten EU-Sprachen übersetzt; je Modell wird ein Mittelwert berichtet. Die MMMLU-Zeile ist Googles eigener mehrsprachiger MMLU-Wert für Gemma 4.
The EU21 numbers are OpenGPT-X’s, taken from the Teuken 7B Instruct v0.6 model card and the accompanying evaluation preprint; they are machine-translated benchmarks, not native-language test sets. The MMMLU number is Google’s own. LLM EU has run no evaluation and has not verified either harness.
| Modelle | score | source |
|---|---|---|
| Teuken 7B Instruct v0.6 | 0.57 | link |
| Llama 3.1 8B Instruct | 0.563 | link |
| Mistral 7B Instruct v0.3 | 0.527 | link |
| Aya 23 8B | 0.485 | link |
| Pharia-1-LLM-7B-control-aligned | 0.417 | link |
| Gemma 4 31B IT | 88.4 | link |