Model
Model explorer
MiMo-V2.5
OPENXiaomi · MiMo family · released Apr 22, 2026
Xiaomi's native omnimodal flagship-tier model (text/image/video/audio), 310B total / 15B active MoE, 1M context, fully open-sourced under MIT. Consolidates the prior separate MiMo-V2-Pro (reasoning) and MiMo-V2-Omni (multimodal) lines into one architecture built on the MiMo-V2-Flash backbone.
ReasoningCodingVisionFunction callingTool useAgentic
2050.2
Elo · rank #75
Parameters
310B
Active params
15B (MoE)
Context
1M tokens
Architecture
Sparse MoE (256 routed experts, 8/token), hybrid Sliding-Window+Global attention 5:1, native 729M-param ViT + 261M-param audio encoder, 3x MTP modules
License
MIT
Languages
—
API price (in/out)
$0.14 / $0.28
Modalities
text · vision · audio · video
Benchmark results
Bar shows position within the tracked field; marker = field best
best: GPT-6 Astra · 59.3%
best: Claude Fable 5 · 1505
best: Qwen3.8-Flash-Next · 90.6%
best: Claude Fable 5 · 1932
best: GPT-6 Astra · 96.0%
best: Claude Opus 5 · 64.7%
best: MiniMax M3 · 83.0%
best: DeepSeek-V4-Pro (Think Max) · 93.5%
best: Claude Fable 5 · 91.5%
best: Claude Opus 4.7 · 85.5%
best: Trinity-Large-Thinking · 91.9%
best: Claude Fable 5.1 · 81.2%
best: Claude Opus 5 · 96.0%
best: Claude Opus 4.6 · 99.3%
best: GPT-5.6 Sol · 33.0%
best: Gemini 3.8 Flash · 89.4%
best: Seed 2.1 Pro · 89.2%
Run it locally
VRAM @ Q4
—
VRAM @ FP16
—
Fits on (Q4)
Multi-node cluster required
Throughput data unavailable.
Quantizations
—
Fine-tune it
PermissiveQLoRA210.8 GB2× H200 141GB
LoRA660.3 GB4× B200 192GB
Full fine-tune4975.5 GBbeyond 8× B200
QLoRA SFT on ~10k samples ≈ $7.76 (2× H200 141GB)
API price $0.14/$0.28 · each benchmark row carries its own source badge (see methodology)