Qwen3-8B
Suverenitāte un mitināšana
Šis modelis ir iekļauts katalogā uzziņai. Tas vēl nav mitināts LLM EU, un neviens pirmās puses galapunkts to šeit neapkalpo. Pieprasījumi tam tiek nosūtīti tieši nodrošinātājam saskaņā ar šī nodrošinātāja nosacījumiem.
Kopsavilkums
Qwen3-8B is a dense text model that supports both a thinking and a non-thinking mode in one checkpoint. It handles 32,768 tokens natively, and the model card states that YaRN scaling extends it to 131,072 tokens. It is published under Apache-2.0 and remains one of the most downloaded open text models in its size class.
Specifikācijas
- Nodrošinātājs
- Alibaba Qwen (CN)
- Konteksta logs
- 33k
- Parametri
- 8.19B
- Modalitātes
- text
- Licence
- apache-2.0 · licence
- Atvērti svari
- yes
- Izlaiduma datums
- 2025. g. 27. apr.
- Uzdevumi
- cheap, reasoning, multilingual-support, summarization
Valodas
Spēcīga: en, zh
Atbilstoša: de, fr, es, it, pt, ru, ja, ko, ar, nl, pl, tr, vi, hi
Valodas
Spēcīga
- en
- zh
Atbilstoša
- de
- fr
- es
- it
- pt
- ru
- ja
- ko
- ar
- nl
- pl
- tr
- vi
- hi
Nav verificēts: nodrošinātājs nav dokumentējis šo apgalvojumu. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.
Kad to izmantot
Stiprās puses
- hybrid thinking and non-thinking modes
- Apache-2.0 licence
- runs on a single mid-range GPU
Ierobežojumi
- 32,768-token native context; longer inputs need YaRN configuration
- behind Qwen3.5 and Qwen3.8 releases on reasoning
- no evaluation has been run by LLM EU