Qwen3 Coder 480B A35B qwen3-coder-480b-a35b
qwen · qwen coding flagship
#32 overall #7 agentic
- params
- 480B (A35B)
- arch
- moe
- context
- 256k
- license
- apache-2.0 open weights
- released
- Jul 2025
- train compute
- 1.6e+24 FLOP
# architecture
- layers
- 62
- d_model
- 6144
- heads
- 96q / 8kv
- head dim
- 128
- experts
- top-8 of 160
- vocab
- 151936
- family
- qwen3_moe
signal path of one layer, generated from the registry's structured fields — dimension lines quote the model's real numbers; an MoE trace forks at the router, a dense trace runs straight through. Geometry fields come from the repo's config.json.
# why ranked
Overall: #32 — score 51.6 ● high — all signals present
| signal | weight | input (0–100) |
|---|---|---|
| aa_intelligence | 0.70 | 42.3 |
| bench_composite | 0.30 | 73.4 |
benchmark panel evidence:
| complete panel | score (0–100) |
|---|---|
| swe-bench-verified-v1 | 73.4 |
Coding: provisional — score 72.2 ● low — a single signal; treat with caution
provisional: too few independent signals for a numbered position — the score is shown, but this model sorts after every ranked model.
| signal | weight | input (0–100) |
|---|---|---|
| aa_coding | 0.30 | — |
| aider_polyglot | 0.30 | — |
| swe_bench_verified | 0.40 | 72.2 |
missing signals are dropped and the remaining weights renormalized — never imputed.
Agentic: #7 — score 72.2 ● low — a single signal; treat with caution
| signal | weight | input (0–100) |
|---|---|---|
| swe_bench_verified | 0.60 | 72.2 |
| swe_rebench | 0.40 | — |
missing signals are dropped and the remaining weights renormalized — never imputed.
# trend
methodology changed during this history window; score movement across that boundary is not model movement. See methodology v5.
# benchmarks
| benchmark | score | source | date |
|---|---|---|---|
| AA Intelligence Index | 11.9 | Artificial Analysis | — |
| AA Math Index | 39.3 | Artificial Analysis | — |
| SWE-bench Verified | 69.6 | SWE-bench | Aug 2025 |
| LMArena Elo (Code) | 1273.0 | LMArena | Aug 2026 |
# where to run
| provider | quant | ctx | $/M in | $/M out | $/M cache | price src | tps | uptime | ✓ |
|---|---|---|---|---|---|---|---|---|---|
| Featherless | — | — | — | — | via hfrouter | — | — | ||
| SiliconFlow | 262k | $0.25 | $1.00 | — | via modelsdev | — | — | ||
| DeepInfra | fp4 | 256k | $0.30 | $1.00 | $0.10 | via openrouter | — | 97.7% | |
| unknown | 256k | $0.22 | $1.80 | — | via openrouter | — | 99.9% | ||
| Venice | fp8 | 250k | $0.35 | $1.50 | $0.04 | via openrouter | — | 89.1% | |
| Novita | fp8 | 256k | $0.38 | $1.55 | — | novita | — | 95.8% | ✓ |
| Amazon Bedrock | — | $0.45 | $1.80 | — | bedrock | — | — | ✓ | |
| Wandb | 256k | $1.00 | $1.50 | — | via litellm | — | — | ||
| Alibaba | unknown | 256k | $0.97 | $4.88 | — | via openrouter | — | 100.0% | |
| Hyperbolic | 256k | $2.00 | $2.00 | — | hyperbolic | — | — | ✓ | |
| Together | 256k | $2.00 | $2.00 | — | via modelsdev | — | — |
sorted by blended price ((3·input + output) / 4 per 1M) · ✓ = the provider's own catalog confirms the offer · "via …" prices are what the aggregator routing the offer charges, not the provider's own list price
source aliases
- aa
qwen3-coder-480b,qwen3-coder-480b-a35b-instruct- arena
Qwen3-Coder-480B-A35B-Instruct- openrouter
qwen/qwen3-coder- swebench
Qwen3-Coder-480B-A35B-Instruct