Model
Model explorer
Qwen3-Coder-480B-A35B
OPENAlibaba · Qwen3-Coder family · released Jul 22, 2025
Most powerful open agentic coding MoE; 256K native (1M extrapolated) context, rivals Claude Sonnet 4.
ReasoningCodingVisionFunction callingTool useAgentic
1326.5
Elo · rank #165
Parameters
480B
Active params
35B (MoE)
Context
256K tokens
Architecture
MoE, 160 experts / 8 active
License
Apache 2.0
Languages
119+
API price (in/out)
$1.5 / $7.5
Modalities
text
Benchmark results
Bar shows position within the tracked field; marker = field best
best: Claude Opus 4.5 · 89.4%
best: Hunyuan-A13B · 78.3%
best: GPT-6 Astra · 96.0%
best: Claude Opus 5 · 64.7%
best: DeepSeek-V4-Pro (Think Max) · 93.5%
best: Claude Fable 5.1 · 81.2%
best: Claude Opus 5 · 96.0%
best: Claude Opus 4.1 · 82.4%
best: Gemini 3.8 Flash · 89.4%
Run it locally
VRAM @ Q4
289 GB
VRAM @ FP16
960 GB
Fits on (Q4)
M3 Ultra 512GB
Far too large for a single RTX 4090; needs multi-GPU or large CPU-RAM offload setups.
Quantizations
GGUF · FP8 · AWQ
Fine-tune it
PermissiveQLoRA326.4 GB2× B200 192GB
LoRA1022.4 GB8× H200 141GB
Full fine-tune7704.0 GBbeyond 8× B200
QLoRA SFT on ~10k samples ≈ $12.52 (2× B200 192GB)
API price $1.5/$7.5 · each benchmark row carries its own source badge (see methodology)