Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Reviews and evidence

MiMo-V2.6-Pro · Media / benchmark · Independent measurement

Artificial Analysis: MiMo-V2.6-Pro Intelligence Index, Speed, Pricing, and Latency

Under the Xiaomi provider, reasoning variant, and default production workload of 10,000 input tokens, Artificial Analysis measured MiMo-V2.6-Pro at an Intelligence Index of 46.324 (displayed as 46), output speed of 129.75 tokens/s, 17.58 seconds to the first answer token, and 21.44 seconds for an end-to-end 500-token output; pricing is in the low-cost range。

Media / benchmarkIndependent measurementEdited 2026-09-22

Test conditions

Source-specific observation
Under the Xiaomi provider, reasoning variant, and default production workload of 10,000 input tokens, Artificial Analysis measured MiMo-V2.6-Pro at an Intelligence Index of 46.324 (displayed as 46), output speed of 129.75 tokens/s, 17.58 seconds to the first answer token, and 21.44 seconds for an end-to-end 500-token output; pricing is in the low-cost range。
Published conditions
Interactive short-form Q&A where time to first response is highly sensitive; rigorous cross-provider, cross-reasoning-level, or cross-random-seed comparisons; this page has only one Xiaomi provider and cannot represent other deployments.

Key data and applicable tasks

One-sentence conclusion

Under the Xiaomi provider, reasoning variant, and default production workload of 10,000 input tokens, Artificial Analysis measured MiMo-V2.6-Pro at an Intelligence Index of 46.324 (displayed as 46), output speed of 129.75 tokens/s, 17.58 seconds to the first answer token, and 21.44 seconds for an end-to-end 500-token output; pricing is in the low-cost range, but reasoning wait time accounts for most of the time to the first answer token.

Suitable use cases

  • Suitable tasks: Tasks requiring strong general reasoning, long context, multimodal input, or Agent workflows while also prioritizing API unit cost and output speed; the AA page lists text, image, audio, and video as input modalities, with a 1M-token context window.

  • Unsuitable tasks: Interactive short-form Q&A where time to first response is highly sensitive; rigorous cross-provider, cross-reasoning-level, or cross-random-seed comparisons; this page has only one Xiaomi provider and cannot represent other deployments.

  • Applicable model version: The reasoning version of MiMo-V2.6-Pro. The results must not be extrapolated to MiMo-V2.6-Pro-RL, MiMo-V2.6-Pro-Ultraspeed, MiMo-V2.6-Flash, or other non-reasoning variants.

  • Applicable client, Agent, or API: Artificial Analysis's independent benchmark environment and the Xiaomi API provider listed on the page; the page does not publish a complete, directly reusable request payload, tool schema, or evaluation harness.

  • Recommended reasoning level and parameters: This source does not publish selectable reasoning levels, temperature, top-p, random seed, or complete API parameters; a retest should fix the reasoning version, Xiaomi provider, and 10,000-input-token workload, while recording each unknown parameter.

Test method / workflow steps

  1. Use the Artificial Analysis model page to verify the identity of MiMo-V2.6-Pro, its reasoning label, and its association with the Xiaomi provider; do not mix results for Pro, Flash, RL, or Ultraspeed.

  2. Use Artificial Analysis Intelligence Index v4.3.2 as the composite intelligence measure. The page says the index contains 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, and AA-LCR v1.1.

  3. Read the structured Dataset data embedded in the model page and record the index, output-token usage, cost per task, total evaluation cost, and speed; the index page's measurementTechnique explicitly says that Artificial Analysis ran the measurement independently on dedicated hardware.

  4. On the Provider Benchmark page, use the default workload of 10,000 input tokens. The page's pricing explanation uses a cache-input-output ratio of 7:2:1 for the blended price; speed is output tokens/s; latency is measured by the first answer token and end-to-end output of 500 tokens.

  5. Break down time to the first answer token according to the page's structured data: reasoning time + input processing time = time to the first answer token. Check end-to-end time as input processing + reasoning + output time for 500 tokens.

  6. When comparing with other models, retain the reasoning labels shown on the page (for example, max, high, and xhigh) and each provider; do not treat the page ranking as a causal conclusion under the same reasoning level.

Original evidence and data

1. Model identity and measurement scope

FieldArtificial Analysis page valueVerification note
ModelMiMo-V2.6-ProThe model-page title and the label in the structured data are consistent
CreatorXiaomiThe page shows Xiaomi and labels it an open weights model
reasoningYesThe page notes that the current page shows the reasoning version of this model
ProviderXiaomiThe Provider page states that there is only 1 provider, Xiaomi
Evaluation hardware scopeArtificial Analysis dedicated hardwareThe model page Dataset's measurementTechnique says this is an independent test
Provider workload10,000 input tokensThe Provider page says the default performance benchmark workload was updated to 10k input tokens
Index versionArtificial Analysis Intelligence Index v4.3.2The page explicitly lists the version and its 10 component evaluations
Page rankingIntelligence #1 / 114This is the model-class ranking in the model-page summary, not the global ranking among all 656 models

2. Intelligence Index, speed, tokens, and cost

MetricOriginal valuePage display / calculated valueConditions and limitations
Intelligence Index46.324206531038346v4.3.2; composed of 10 evaluations; not a Xiaomi-reported score
Output speed129.748797446312 tokens/s129.7 / 130 tokens/sXiaomi provider; the page defines this as output tokens per second
Reasoning tokens per index task37,519.68181980028Not fully displayed on the pagereasoning field in the structured Dataset; weighted-average scope across the task set
Answer tokens per index task26,756.018691867608Not fully displayed on the pageanswer field in the structured Dataset
Total output tokens per index task64,275.70051166789Approximately 140M in the page summary for the entire index evaluationreasoning + answer; the page does not provide sample-count breakdowns for each sub-evaluation
Cost per Intelligence Index task0.13322318937213493 USD$0.13The page defines this as the weighted average cost per Intelligence Index task
Cost to run the complete Intelligence Index206.65529999999998 USD$206.66Sum of the structured cost components; not the cost of a single user request
Input price0.435 USD / 1M tokens$0.435Xiaomi provider; uncached input
Output price0.87 USD / 1M tokens$0.87Xiaomi provider
Cache-hit price0.0036 USD / 1M tokensThe page displays <0.01Xiaomi provider
Cache discount99.17241379310345%Approximately 99% in the page summaryCalculated as 1 - cache hit price / input price
Blended price0.17652000000000001 USD / 1M tokens$0.18The Provider page's 7:2:1 cache-input-output blended basis

3. Latency and end-to-end response

Artificial Analysis's latency data explicitly separates reasoning time from input processing time. Its model-page Dataset gives the following MiMo-V2.6-Pro data:

MetricOriginal value (seconds)Page display / calculated valueDefinition
Reasoning time15.41440105313938115.41Thinking time before the first answer token for a reasoning model
Input processing time2.16728361549998022.17Input processing time after the API request; corresponding to Median First Chunk in the Provider table
First answer token17.5816846686393617.5815.414401053139381 + 2.1672836154999802; highlighted on the page as “Time to First Answer Token”
Output time for 500 tokens3.85360026328484523.85Output phase for 500 tokens calculated from measured output speed
End-to-end response time21.43528493192420521.44Input processing + reasoning + output for 500 tokens

Under this measurement scope, reasoning wait time accounts for approximately 15.4144 / 17.5817 ≈ 87.7% of the time to the first answer token; this is a calculation from the page's published components and does not represent other request lengths or providers.

4. Model specifications (for interpreting the measurement scope)

  • Context window: 1,000,000 tokens.

  • Total parameters: approximately 1.0T on the page; the structured model-size Dataset gives 978B passive parameters.

  • Active parameters: 42B.

  • Input: text, image, audio, and video.

  • Output: text.

  • License: MIT.

Scope and limitations

  • This is an independent Artificial Analysis benchmark result, not a Xiaomi-published benchmark and not a local retest; it cannot support a claim that the same speed will be achieved across all APIs, regions, networks, or loads.

  • The page's structured data says that the measurement ran independently on Artificial Analysis dedicated hardware, but it does not publish the complete hardware model, request payload, concurrency, network path, random seed, temperature, top-p, timeout policy, or failed samples. Reproduction can therefore only approximately verify the metrics and workload level.

  • The model-page body lists 10 component evaluations for the Intelligence Index, but does not provide MiMo-V2.6-Pro's score for each item, per-item sample counts, confidence intervals, or per-question logs. The 46.324 is suitable as an aggregate comparison signal and should not be decomposed into conclusions about individual capabilities.

  • The Provider page has only one provider, Xiaomi. The $0.13 per task, $0.18 blended price, and 17.58-second first-answer latency cannot be extrapolated to unlisted providers.

  • The page's prices are public rates per million tokens; the $206.66 cost of the complete Intelligence Index is Artificial Analysis's evaluation expenditure and should not be treated as the price of an ordinary user session.

  • The model-page #1 / 114 and the chart's “28 of 656 models” belong to different filtering/display contexts; they must not be combined into a single ranking.

  • Artificial Analysis's X account, https://x.com/ArtificialAnlys, was opened through Tabbit, but the timeline returned “It looks like your connection is lost. We’ll keep trying to reconnect.” No post directly related to MiMo-V2.6-Pro was visible, so no data was added from its title, summary, or a paraphrase.

Source excerpts or observations (short excerpts for compliance only)

  • The model-page Dataset's public description of the measurement method: Independent test run by Artificial Analysis on dedicated hardware.

  • The Provider page notes that the default performance benchmark workload was updated to 10k input tokens.

  • The Provider page's structured data also gives Xiaomi's 129.74879744631201 output tokens/s, 15.414401053139381 seconds of reasoning time, and 2.1672836154999802 seconds of input processing time; this piece preserves these original fields to sufficient precision and lists the rounded values separately as page-display values.

What this supports

  • Tasks requiring strong general reasoning, long context, multimodal input, or Agent workflows while also prioritizing API unit cost and output speed; the AA page lists text, image, audio, and video as input modalities, with a 1M-token context window.

What this does not support

  • Interactive short-form Q&A where time to first response is highly sensitive; rigorous cross-provider, cross-reasoning-level, or cross-random-seed comparisons; this page has only one Xiaomi provider and cannot represent other deployments.

Method, limits, and reproduction

The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.

For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.

Original source

Artificial Analysis · Artificial Analysis · Original publication date 2026-09-22 · Site edit date 2026-09-22

Open original source

MiMo-V2.6-Pro

Compare MiMo-V2.6-Pro in Tabbit

Download the Tabbit client to check model access

Read the full analysis

Full review · English

MiMo-V2.6-Pro Review: The Smartest Open Model Makes You Wait

A public-evidence review of MiMo-V2.6-Pro: what it does well, where it bites, real user reports, and a workload verdict on Xiaomi's open flagship.

Pricing · English

MiMo-V2.6-Pro Pricing: Official Rate Card, Cache Levers, and Cost per Task

A practical decision guide to MiMo-V2.6-Pro pricing: official API rates, prompt cache economics, reasoning token overhead, UltraSpeed mode, and worked task budgets.

Alternatives · English

MiMo-V2.6-Pro Alternatives: Choose by Task and Budget

Compare five MiMo-V2.6-Pro alternatives by completed-task cost, agentic reliability, open weights, and deployment fit, with prices checked on September 22, 2026.

Comparison · English

MiMo-V2.6-Pro vs MiMo-V2.6-Flash: Which Xiaomi MoE Model Fits Your Workload?

A head-to-head comparison of MiMo-V2.6-Pro and Flash: 1.02T vs 309B MoE architecture, 3.1x pricing delta, reasoning token overhead, agent benchmarks, and decision matrix.

Related reviews

Xiaomi MiMo Official Release: MiMo-V2.6-Pro Benchmark Signals and Native Omnimodal PositioningXiaomi positions MiMo-V2.6-Pro as a native omnimodal open-source model for agents, coding, vision, and computer use, and reports an Artificial Analysis Intelligence Index score of 46 along with several RL/agent results; these figures remain the vendor's own reporting and cannot replace an independent rerun under the same harness.MiMo-V2.6-Pro Official Technical Report: Architecture, Scaled RL, and Evaluation ConditionsThe official report defines MiMo-V2.6-Pro as a native multimodal sparse MoE with 1.02T total parameters and approximately 42B active parameters, and reports strong agent benchmark results from large-batch, multi-environment RL training with multiple harnesses and groupwise graders; however, most figures in the tables are vendor-reported。MiMo-V2.6-Pro Official X Release Thread: Task Positioning, Public Benchmarks, and Open-Source Entry PointsThis official thread positions MiMo-V2.6-Pro as an openly built native omnimodal agent model focused on coding, computer use, 3D, design, research, and tool workflows; its rankings and examples help identify promising task directions, but they remain vendor-reported and cannot replace an independent rerun under the same harness.Arena Code Arena: MiMo-V2.6-Pro WebDev AutoEval RecordArena's Code Arena | WebDev overall leaderboard includes mimo-v2.6-pro with an AutoEval score of 1628 (+18/-18), but it does not publish a vote count or rank, so this only shows that it was included in the WebDev automated evaluation leaderboard; 1628 must not be treated as a blind-test ranking.Xiaomi MiMo-V2.6-Pro Official API Integration and Reasoning ConfigurationThis official configuration can be used to connect to mimo-v2.6-pro through the OpenAI-compatible protocol, with deep thinking, streaming output, and multi-turn tool calls enabled as needed.Hugging Face Official MiMo-V2.6-Pro-RL Local Deployment and Chat Template ConfigurationThe official model card provides SGLang and vLLM service commands for MiMo-V2.6-Pro-RL and defines chat-template behavior for text, image, video, audio, thinking, and tool calls in the repository tokenizer configuration; local deployment must use the checkpoint name and must not treat it as the same model identifier as the hosted API's mimo-v2.6-pro.Xiaomi MiMo-V2.6-Pro Omnimodal Input and Visual Task Workflowmimo-v2.6-pro can read publicly accessible URLs or properly formatted Base64 images, videos, and audio through the OpenAI Chat Completions API, but it cannot directly upload local files, and the combined media and text tokens remain subject to the 1M context limit.Xiaomi MiMo-V2.6-Pro Official Function Calling and Multi-Turn Agent WorkflowFor mimo-v2.6-pro, the official workflow is “the model returns a complete assistant message, including reasoning_content and tool_calls → the client executes the tools → appends the role: tool results → requests the model again,” repeating until the current turn produces no more tool calls.