Parameters
1T (Estimated)
Context Length
1.05M
Modality
Multimodal
Architecture
Undisclosed
License
Proprietary
Release Date
21 Sept 2026
Knowledge Cutoff
-
API Pricing (per 1M)
Input: $4.35 · Output: $8.70
No evaluation benchmarks for MiMo-V2.6-Pro-UltraSpeed available.
Overall Rank
-
Coding Rank
-
MiMo-V2.6-Pro-UltraSpeed is an optimized high-throughput edition of Xiaomi's flagship foundation model designed to match original output quality at substantially lower latency. It excels in complex multimodal tasks, agentic reasoning, and long-context processing.
Architecture specifications are undisclosed for proprietary models.
MiMo-V2-Flash is a Mixture-of-Experts (MoE) model with hybrid attention architecture designed for high-speed reasoning and agentic workflows. It features Multi-Token Prediction (MTP) to achieve state-of-the-art performance while significantly reducing inference costs. The model is optimized for long-context modeling and efficient inference.
Assistant
Online