MiMo-V2.6-Pro · Media / benchmark · Independent measurement
Under the Xiaomi provider, reasoning variant, and default production workload of 10,000 input tokens, Artificial Analysis measured MiMo-V2.6-Pro at an Intelligence Index of 46.324 (displayed as 46), output speed of 129.75 tokens/s, 17.58 seconds to the first answer token, and 21.44 seconds for an end-to-end 500-token output; pricing is in the low-cost range。
Under the Xiaomi provider, reasoning variant, and default production workload of 10,000 input tokens, Artificial Analysis measured MiMo-V2.6-Pro at an Intelligence Index of 46.324 (displayed as 46), output speed of 129.75 tokens/s, 17.58 seconds to the first answer token, and 21.44 seconds for an end-to-end 500-token output; pricing is in the low-cost range, but reasoning wait time accounts for most of the time to the first answer token.
Suitable tasks: Tasks requiring strong general reasoning, long context, multimodal input, or Agent workflows while also prioritizing API unit cost and output speed; the AA page lists text, image, audio, and video as input modalities, with a 1M-token context window.
Unsuitable tasks: Interactive short-form Q&A where time to first response is highly sensitive; rigorous cross-provider, cross-reasoning-level, or cross-random-seed comparisons; this page has only one Xiaomi provider and cannot represent other deployments.
Applicable model version: The reasoning version of MiMo-V2.6-Pro. The results must not be extrapolated to MiMo-V2.6-Pro-RL, MiMo-V2.6-Pro-Ultraspeed, MiMo-V2.6-Flash, or other non-reasoning variants.
Applicable client, Agent, or API: Artificial Analysis's independent benchmark environment and the Xiaomi API provider listed on the page; the page does not publish a complete, directly reusable request payload, tool schema, or evaluation harness.
Recommended reasoning level and parameters: This source does not publish selectable reasoning levels, temperature, top-p, random seed, or complete API parameters; a retest should fix the reasoning version, Xiaomi provider, and 10,000-input-token workload, while recording each unknown parameter.
Use the Artificial Analysis model page to verify the identity of MiMo-V2.6-Pro, its reasoning label, and its association with the Xiaomi provider; do not mix results for Pro, Flash, RL, or Ultraspeed.
Use Artificial Analysis Intelligence Index v4.3.2 as the composite intelligence measure. The page says the index contains 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, and AA-LCR v1.1.
Read the structured Dataset data embedded in the model page and record the index, output-token usage, cost per task, total evaluation cost, and speed; the index page's measurementTechnique explicitly says that Artificial Analysis ran the measurement independently on dedicated hardware.
On the Provider Benchmark page, use the default workload of 10,000 input tokens. The page's pricing explanation uses a cache-input-output ratio of 7:2:1 for the blended price; speed is output tokens/s; latency is measured by the first answer token and end-to-end output of 500 tokens.
Break down time to the first answer token according to the page's structured data: reasoning time + input processing time = time to the first answer token. Check end-to-end time as input processing + reasoning + output time for 500 tokens.
When comparing with other models, retain the reasoning labels shown on the page (for example, max, high, and xhigh) and each provider; do not treat the page ranking as a causal conclusion under the same reasoning level.
| Field | Artificial Analysis page value | Verification note |
|---|---|---|
| Model | MiMo-V2.6-Pro | The model-page title and the label in the structured data are consistent |
| Creator | Xiaomi | The page shows Xiaomi and labels it an open weights model |
| reasoning | Yes | The page notes that the current page shows the reasoning version of this model |
| Provider | Xiaomi | The Provider page states that there is only 1 provider, Xiaomi |
| Evaluation hardware scope | Artificial Analysis dedicated hardware | The model page Dataset's measurementTechnique says this is an independent test |
| Provider workload | 10,000 input tokens | The Provider page says the default performance benchmark workload was updated to 10k input tokens |
| Index version | Artificial Analysis Intelligence Index v4.3.2 | The page explicitly lists the version and its 10 component evaluations |
| Page ranking | Intelligence #1 / 114 | This is the model-class ranking in the model-page summary, not the global ranking among all 656 models |
| Metric | Original value | Page display / calculated value | Conditions and limitations |
|---|---|---|---|
| Intelligence Index | 46.3242065310383 | 46 | v4.3.2; composed of 10 evaluations; not a Xiaomi-reported score |
| Output speed | 129.748797446312 tokens/s | 129.7 / 130 tokens/s | Xiaomi provider; the page defines this as output tokens per second |
| Reasoning tokens per index task | 37,519.68181980028 | Not fully displayed on the page | reasoning field in the structured Dataset; weighted-average scope across the task set |
| Answer tokens per index task | 26,756.018691867608 | Not fully displayed on the page | answer field in the structured Dataset |
| Total output tokens per index task | 64,275.70051166789 | Approximately 140M in the page summary for the entire index evaluation | reasoning + answer; the page does not provide sample-count breakdowns for each sub-evaluation |
| Cost per Intelligence Index task | 0.13322318937213493 USD | $0.13 | The page defines this as the weighted average cost per Intelligence Index task |
| Cost to run the complete Intelligence Index | 206.65529999999998 USD | $206.66 | Sum of the structured cost components; not the cost of a single user request |
| Input price | 0.435 USD / 1M tokens | $0.435 | Xiaomi provider; uncached input |
| Output price | 0.87 USD / 1M tokens | $0.87 | Xiaomi provider |
| Cache-hit price | 0.0036 USD / 1M tokens | The page displays <0.01 | Xiaomi provider |
| Cache discount | 99.17241379310345% | Approximately 99% in the page summary | Calculated as 1 - cache hit price / input price |
| Blended price | 0.17652000000000001 USD / 1M tokens | $0.18 | The Provider page's 7:2:1 cache-input-output blended basis |
Artificial Analysis's latency data explicitly separates reasoning time from input processing time. Its model-page Dataset gives the following MiMo-V2.6-Pro data:
| Metric | Original value (seconds) | Page display / calculated value | Definition |
|---|---|---|---|
| Reasoning time | 15.414401053139381 | 15.41 | Thinking time before the first answer token for a reasoning model |
| Input processing time | 2.1672836154999802 | 2.17 | Input processing time after the API request; corresponding to Median First Chunk in the Provider table |
| First answer token | 17.58168466863936 | 17.58 | 15.414401053139381 + 2.1672836154999802; highlighted on the page as “Time to First Answer Token” |
| Output time for 500 tokens | 3.8536002632848452 | 3.85 | Output phase for 500 tokens calculated from measured output speed |
| End-to-end response time | 21.435284931924205 | 21.44 | Input processing + reasoning + output for 500 tokens |
Under this measurement scope, reasoning wait time accounts for approximately 15.4144 / 17.5817 ≈ 87.7% of the time to the first answer token; this is a calculation from the page's published components and does not represent other request lengths or providers.
Context window: 1,000,000 tokens.
Total parameters: approximately 1.0T on the page; the structured model-size Dataset gives 978B passive parameters.
Active parameters: 42B.
Input: text, image, audio, and video.
Output: text.
License: MIT.
This is an independent Artificial Analysis benchmark result, not a Xiaomi-published benchmark and not a local retest; it cannot support a claim that the same speed will be achieved across all APIs, regions, networks, or loads.
The page's structured data says that the measurement ran independently on Artificial Analysis dedicated hardware, but it does not publish the complete hardware model, request payload, concurrency, network path, random seed, temperature, top-p, timeout policy, or failed samples. Reproduction can therefore only approximately verify the metrics and workload level.
The model-page body lists 10 component evaluations for the Intelligence Index, but does not provide MiMo-V2.6-Pro's score for each item, per-item sample counts, confidence intervals, or per-question logs. The 46.324 is suitable as an aggregate comparison signal and should not be decomposed into conclusions about individual capabilities.
The Provider page has only one provider, Xiaomi. The $0.13 per task, $0.18 blended price, and 17.58-second first-answer latency cannot be extrapolated to unlisted providers.
The page's prices are public rates per million tokens; the $206.66 cost of the complete Intelligence Index is Artificial Analysis's evaluation expenditure and should not be treated as the price of an ordinary user session.
The model-page #1 / 114 and the chart's “28 of 656 models” belong to different filtering/display contexts; they must not be combined into a single ranking.
Artificial Analysis's X account, https://x.com/ArtificialAnlys, was opened through Tabbit, but the timeline returned “It looks like your connection is lost. We’ll keep trying to reconnect.” No post directly related to MiMo-V2.6-Pro was visible, so no data was added from its title, summary, or a paraphrase.
The model-page Dataset's public description of the measurement method: Independent test run by Artificial Analysis on dedicated hardware.
The Provider page notes that the default performance benchmark workload was updated to 10k input tokens.
The Provider page's structured data also gives Xiaomi's 129.74879744631201 output tokens/s, 15.414401053139381 seconds of reasoning time, and 2.1672836154999802 seconds of input processing time; this piece preserves these original fields to sufficient precision and lists the rounded values separately as page-display values.
The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.
For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.
Artificial Analysis · Artificial Analysis · Original publication date 2026-09-22 · Site edit date 2026-09-22
Open original sourceMiMo-V2.6-Pro
Download the Tabbit client to check model access