Creator
Xiaomi
Released
2026-04-22
Intelligence
42.2
Artificial Analysis Index
Coding
60.2
Artificial Analysis Index
In $/1M
$0.43
input tokens
Out $/1M
$0.87
output tokens
Blended $/1M
$0.54
3:1 blended
Speed
63
tokens / sec
Profile
License, openness, and modality — the governance layer.
License
MIT
commercial OK
Weights
Open
downloadable
Modalities
Text
Parameters
1.0T
total
MiMo-V2.5-Pro is Xiaomi's 1.02T-parameter sparse Mixture-of-Experts language model with 42B active parameters and a 1M-token context window. It inherits the MiMo-V2-Flash hybrid-attention and Multi-Token Prediction design, extends context during pre-training up to 1M tokens, and uses supervised fine-tuning, domain-specialized reinforcement learning, and Multi-Teacher On-Policy Distillation to improve complex software engineering, long-horizon agentic tasks, and ultra-long-context coherence.
Capability profile
Category strength across 20 domains, via LLM Stats.
Benchmark breakdown
Independent evaluation scores, normalized to 0–100.
Providers
Inference hosts serving this model — their own pricing and measured performance, cheapest input first.
| Provider | In $/1M | Out $/1M | Throughput | Latency | Status |
|---|---|---|---|---|---|
| Xiaomi | $0.43 | $0.87 | — | — | active |
| DeepInfra | $1.00 | $3.00 | — | — | active |
| Novita | $2.00 | $6.00 | — | — | active |
3 hosts · via LLM Stats