DataLearner logo
MI

MiMo-V2.6-Pro-UltraSpeed

Multimodal modelCoding model

MiMo-V2.6-Pro-UltraSpeed

Release date: 2026-09-22Views: 1
Live demoGitHubHugging FaceCompare
Parameters
1T
Context length
1M
Multilingual
Supported
Reasoning ability
5/5

MiMo-V2.6-Pro-UltraSpeed is the Xiaomi MiMo platform's high-speed serving tier for MiMo-V2.6-Pro, offering up to 20x output speed at roughly 10x the standard price.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

MiMo-V2.6-Pro-UltraSpeed

Model basics

Reasoning traces
Supported
Thinking modes
Thinking modes not supported
Context length
1M tokens
Max output length
No data
Model type
Multimodal model
Modality (in / out)
Text, Image, Audio, Video → Text
Release date
2026-09-22
Model file size
No data
MoE architecture
Yes
Total params / Active params
1T / 42B
Knowledge cutoff
No data
MiMo-V2.6-Pro-UltraSpeed

Open source & experience

Code license
Weights license
MIT License- Commercial use permitted
GitHub repo
N/A
Live demo
N/A
MiMo-V2.6-Pro-UltraSpeed

Official resources

Paper
DataLearnerAI blog
N/A
MiMo-V2.6-Pro-UltraSpeed

API details

API speed
5/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-¥30.00/ 1M tokens¥60.00/ 1M tokens
international
TypeConditionInputOutput
Text-$4.35/ 1M tokens$8.70/ 1M tokens
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$0.036/ 1M tokens
Text-¥0.250/ 1M tokens

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

MiMo-V2.6-Pro-UltraSpeed

Publisher

MiMo-V2.6-Pro-UltraSpeed

Model Overview

MiMo-V2.6-Pro-UltraSpeed is the high-speed serving tier the Xiaomi MiMo platform offers for MiMo-V2.6-Pro. Xiaomi says it keeps V2.6-Pro's capability while delivering up to 20x output speed, aimed at real-time interaction and latency-sensitive production workloads. Xiaomi has not documented how the tier differs technically from the standard tier, and has not published separate benchmark results for it.


Specifications

Capabilities match MiMo-V2.6-Pro: a sparse MoE with 1.02 trillion total parameters and 42B activated. The models are natively omni-modal: they accept text, image, video and audio input and return text, with a 1M-token context window. The vision encoder is a 681M-parameter MiMo ViT (28 layers, 24 SWA + 4 full attention); the audio stack combines a 308M AudioTokenizer with a 127M audio patch encoder, and a 5-layer MTP speculative decoder predicts 7 tokens per pass.


Pricing

The tier is priced separately: ¥30 per million input tokens, ¥0.25 per million cached input tokens and ¥60 per million output tokens domestically; internationally $4.35 input, $0.036 cached input and $8.7 output per million tokens — roughly 10x the standard tier.


Release and license

The MiMo-V2.6 series was published on Hugging Face on September 21, 2026 under the MIT license, and served through the Xiaomi MiMo platform (mimo.mi.com) from September 22, with distribution also via ModelScope and OpenRouter. Xiaomi published a live dashboard of the RL post-training run: Pro and Flash each took under six days and 30 steps, about 750,000 trajectories in total, at a reported cost of roughly $2.62M and $0.85M respectively.

MiMo-V2.6-Pro-UltraSpeed

FAQ

What is MiMo-V2.6-Pro-UltraSpeed?

It is a high-speed serving tier of MiMo-V2.6-Pro on the Xiaomi MiMo platform, advertised at up to 20x output speed.

How much does UltraSpeed cost?

International pricing is $4.35 per million input tokens, $0.036 per million cached input tokens and $8.7 per million output tokens, about 10x the standard MiMo-V2.6-Pro tier.

Does UltraSpeed have its own benchmark scores?

No. Xiaomi has not published separate evaluation results for the UltraSpeed tier; it is presented as the same MiMo-V2.6-Pro capability served faster.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code