GLM-4.7-Flash
Sovereignty and hosting
This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.
Summary
GLM-4.7-Flash is a 30B-A3B mixture-of-experts model with 31.2B parameters, released under the MIT licence with a configuration maximum of 202,752 tokens. The model card reports 91.6 on AIME 25 and 59.2 on SWE-bench Verified, with a comparison table against Qwen3-30B-A3B-Thinking and gpt-oss-20b. It is positioned as a lightweight deployment option in the 30B class.
Specifications
- Provider
- Z.ai (Zhipu AI) (CN)
- Context window
- 203k
- Parameters
- 31.22B
- Modalities
- text
- License
- mit · licence
- Open weights
- yes
- Release date
- 19 Jan 2026
- Tasks
- reasoning, coding, cheap, summarization
Languages
Strong: en, zh
Adequate: de, fr, es, it, pt, ru, ja, ko
Languages
Strong
- en
- zh
Adequate
- de
- fr
- es
- it
- pt
- ru
- ja
- ko
Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.
When to use it
Strengths
- MIT licence with a permissive commercial position
- high published maths score for a 30B-class model
- roughly 200K-token context window
Limits
- benchmark table is self-reported by the vendor
- English and Chinese are the only declared languages
- no evaluation has been run by LLM EU