Model
Model explorer

Qwen3-Coder-480B-A35B

OPEN
Alibaba · Qwen3-Coder family · released Jul 22, 2025

Most powerful open agentic coding MoE; 256K native (1M extrapolated) context, rivals Claude Sonnet 4.

ReasoningCodingVisionFunction callingTool useAgentic
1326.5
Elo · rank #165
Parameters
480B
Active params
35B (MoE)
Context
256K tokens
Architecture
MoE, 160 experts / 8 active
License
Apache 2.0
Languages
119+
API price (in/out)
$1.5 / $7.5
Modalities
text
Benchmark results
Bar shows position within the tracked field; marker = field best
Aider PolyglotCoding61.8%#17
best: Claude Opus 4.5 · 89.4%
BFCL v3Agents68.7%#12
best: Hunyuan-A13B · 78.3%
GPQA DiamondReasoning62.0%#166
best: GPT-6 Astra · 96.0%
Humanity's Last ExamReasoning4.0%#141
best: Claude Opus 5 · 64.7%
LiveCodeBenchCoding44.9%#127
best: DeepSeek-V4-Pro (Think Max) · 93.5%
SWE-bench ProCoding38.7%#54
best: Claude Fable 5.1 · 81.2%
SWE-bench VerifiedCoding66.5%#70
best: Claude Opus 5 · 96.0%
tau-benchAgents60.0%#15
best: Claude Opus 4.1 · 82.4%
Terminal-Bench 2.0Coding23.9%#71
best: Gemini 3.8 Flash · 89.4%
Run it locally
VRAM @ Q4
289 GB
VRAM @ FP16
960 GB
Fits on (Q4)
M3 Ultra 512GB
Far too large for a single RTX 4090; needs multi-GPU or large CPU-RAM offload setups.
Quantizations
GGUF · FP8 · AWQ
Fine-tune it
Permissive
QLoRA326.4 GB2× B200 192GB
LoRA1022.4 GB8× H200 141GB
Full fine-tune7704.0 GBbeyond 8× B200
QLoRA SFT on ~10k samples ≈ $12.52 (2× B200 192GB)
Qwen3-Coder family
Elo progression across releases
API price $1.5/$7.5 · each benchmark row carries its own source badge (see methodology)