GLM-5 glm-5
zai-org · glm hybrid-reasoning flagship agentic
#9 overall #3 agentic
- params
- 754B
- arch
- moe
- context
- 200k
- license
- mit open weights
- released
- Feb 2026
- reasoning
- yes
- train compute
- 6.8e+24 FLOP
# architecture
- layers
- 78
- d_model
- 6144
- heads
- 64q / 64kv
- head dim
- 64
- experts
- top-8 of 256 + 1 shared
- vocab
- 154880
- family
- glm_moe_dsa
signal path of one layer, generated from the registry's structured fields — dimension lines quote the model's real numbers; an MoE trace forks at the router, a dense trace runs straight through. Geometry fields come from the repo's config.json.
# why ranked
Overall: #9 — score 86.3 ● high — all signals present
| signal | weight | input (0–100) |
|---|---|---|
| aa_intelligence | 0.70 | 84.5 |
| bench_composite | 0.30 | 90.5 |
benchmark panel evidence:
| complete panel | score (0–100) |
|---|---|
| browsecomp-v1 | 90.9 |
| swe-bench-verified-v1 | 90.0 |
Coding: provisional — score 94.4 ● low — a single signal; treat with caution
provisional: too few independent signals for a numbered position — the score is shown, but this model sorts after every ranked model.
| signal | weight | input (0–100) |
|---|---|---|
| aa_coding | 0.30 | — |
| aider_polyglot | 0.30 | — |
| swe_bench_verified | 0.40 | 94.4 |
missing signals are dropped and the remaining weights renormalized — never imputed.
Agentic: #3 — score 94.4 ● low — a single signal; treat with caution
| signal | weight | input (0–100) |
|---|---|---|
| swe_bench_verified | 0.60 | 94.4 |
| swe_rebench | 0.40 | — |
missing signals are dropped and the remaining weights renormalized — never imputed.
# trend
methodology changed during this history window; score movement across that boundary is not model movement. See methodology v5.
# benchmarks
| benchmark | score | source | date |
|---|---|---|---|
| AA Intelligence Index | 32.4 | Artificial Analysis | — |
| Browsecomp | 0.8 / 1 | LLM Stats | — |
| SWE-bench Verified | 72.8 | SWE-bench | Feb 2026 |
| LMArena Elo (Code) | 1436.0 | LMArena | Aug 2026 |
# where to run
| provider | quant | ctx | $/M in | $/M out | $/M cache | price src | tps | uptime | ✓ |
|---|---|---|---|---|---|---|---|---|---|
| Featherless | — | — | — | — | via hfrouter | — | — | ||
| GMICloud | fp8 | 198k | $0.60 | $1.92 | $0.12 | via openrouter | — | 99.6% | |
| StreamLake | fp8 | 198k | $0.60 | $1.92 | $0.12 | via openrouter | — | 99.9% | |
| DeepInfra | fp4 | 198k | $0.60 | $2.08 | $0.12 | via openrouter | — | 99.8% | |
| Baidu | fp8 | 198k | $0.70 | $2.24 | $0.14 | via openrouter | — | 100.0% | |
| SiliconFlow | fp8 | 200k | $0.95 | $2.55 | $0.20 | openrouter | — | 99.9% | ✓ |
| AtlasCloud | fp8 | 198k | $0.95 | $3.15 | — | atlascloud | — | — | ✓ |
| BaseTen | 202k | $0.95 | $3.15 | $0.20 | via modelsdev | — | — | ||
| Amazon Bedrock | unknown | 198k | $1.00 | $3.20 | — | bedrock | — | 95.9% | ✓ |
| DigitalOcean | unknown | 64k | $1.00 | $3.20 | $0.20 | via openrouter | — | 93.0% | |
| Together | 198k | $1.00 | $3.20 | — | via modelsdev | — | — | ||
| Venice | fp8 | 198k | $1.00 | $3.20 | — | venice | — | 85.2% | ✓ |
| Z.AI | fp8 | 198k | $1.00 | $3.20 | $0.20 | via openrouter | — | 100.0% | |
| Novita | fp8 | 202k | $1.20 | $4.00 | — | novita | — | 100.0% | ✓ |
sorted by blended price ((3·input + output) / 4 per 1M) · ✓ = the provider's own catalog confirms the offer · "via …" prices are what the aggregator routing the offer charges, not the provider's own list price
# variants
source aliases
- aa
glm-5- epoch
GLM-5- openrouter
z-ai/glm-5