openai
// xiaomi
MiMo-V2.6-Pro-UltraSpeed
xiaomi/mimo-v2.6-pro-ultraspeed
MiMo-V2.6-Pro-UltraSpeed costs $4.35 per 1M input tokens and $8.70 per 1M output tokens with a 1.0M-token context window.
A reasoning, tool-calling model from xiaomi, released Sep 21, 2026. LiveBench hasn't published a run for it, so this page carries measured price and specs only — no benchmark numbers.
- Input / 1M
- $4.35
- Output / 1M
- $8.70
- Context
- 1.0M
- LiveBench overall
- Not evaluated
Specs and pricing
| Input price USD per 1M prompt tokens | $4.35 |
|---|---|
| Output price USD per 1M completion tokens, reasoning included | $8.70 |
| Blended price 3:1 input:output — #35 most expensive of 343 listed models | $5.44 |
| Cached input read Blank means not priced separately — not free | $0.036 |
| Cache write | — |
| Context window | 1,048,576 tokens |
| Max output | 131,072 tokens |
| Input modalities | text, image, video, audio |
| Tool calling | Yes |
| Extended reasoning | Yes |
| Knowledge cutoff | — |
| Listed | Sep 21, 2026 |
What MiMo-V2.6-Pro-UltraSpeed costs to run
Monthly list cost across five workload shapes, using the published cached-input rate where there is one. A floor, not a quote — batch discounts and cache writes aren't included.
| Workload | Per month |
|---|---|
| Support chatbot 1.2K in / 400 out × 200K requests | $1,429/mo |
| RAG assistant 8K in / 600 out × 100K requests | $2,276/mo |
| Coding agent 40K in / 4K out × 20K requests | $1,760/mo |
| Document extraction 20K in / 1.5K out × 50K requests | $4,787/mo |
| Bulk classification 500 in / 20 out × 5M requests | $9,588/mo |
Compare MiMo-V2.6-Pro-UltraSpeed with
Benchmarked models at a similar price.
moonshotai
MiMo-V2.6-Pro-UltraSpeed vs Kimi K3
anthropic
MiMo-V2.6-Pro-UltraSpeed vs Claude Sonnet 4.6
openai
MiMo-V2.6-Pro-UltraSpeed vs GPT-5.2-Codex
openai
MiMo-V2.6-Pro-UltraSpeed vs GPT-5.2
openai
MiMo-V2.6-Pro-UltraSpeed vs GPT-5.6 Terra
MiMo-V2.6-Pro-UltraSpeed FAQ
How much does MiMo-V2.6-Pro-UltraSpeed cost?
MiMo-V2.6-Pro-UltraSpeed lists at $4.35 per 1M input tokens and $8.70 per 1M output tokens, with cached input reads at $0.036 per 1M. A RAG assistant handling 100K requests a month (8K tokens in, 600 out) comes to about $2,276 at list price.
What is the context window of MiMo-V2.6-Pro-UltraSpeed?
MiMo-V2.6-Pro-UltraSpeed accepts up to 1,048,576 tokens of context and can return up to 131,072 output tokens per response.
Does MiMo-V2.6-Pro-UltraSpeed support tool calling?
Yes — MiMo-V2.6-Pro-UltraSpeed supports tool (function) calling and exposes extended reasoning. It accepts text, image, video, audio input.
How does MiMo-V2.6-Pro-UltraSpeed score on benchmarks?
LiveBench has not published a run for MiMo-V2.6-Pro-UltraSpeed yet. Until it does, compare it on price and specs, or run your own evaluation on your own tasks.
What is the API model ID for MiMo-V2.6-Pro-UltraSpeed?
On OpenRouter the model ID is "xiaomi/mimo-v2.6-pro-ultraspeed". It was listed on Sep 21, 2026.
More from xiaomi
How these numbers are produced
- Price and specs — provider list data from OpenRouter, refreshed every 15 minutes. “Blended” is a 3:1 input:output mix.
- Scores and ranks — LiveBench release 2026-06-25. Ranks count only models that have a published run; a blank means “not evaluated”, never “bad”.
- Cost per point — the measured dollars LiveBench spent on the run, divided by the score it earned.
Published benchmarks rank models on someone else's tasks. Before committing, see LLM & agent evaluation for building an eval on your own.