KIE summarizes Grok 4.6's Intelligence Index as 61, tied with GPT-5.6 Sol; its price is $2 for input and $6 for output per 1M tokens. The article argues that its main selling point is its price-to-intelligence ratio, rather than leading on every individual evaluation.
Grok 4.6's strengths are concentrated in knowledge work and legal reasoning: GDPVal-AA 1753, AA-Briefcase 1577, and Harvey LAB 15.8%.
Its Terminal-Bench v3.0 score is 26%, below GPT-5.6 Sol's 34.6%, making it one of the clearest weaknesses in the release table.
Its DeepSWE v1.1 score is 65.9%, a substantial increase from Grok 4.5's 54%, but still below Sol's 73%.
APEX-Agents rose from 47.1% to 57.5%, indicating clear progress on long-horizon Agent tasks.
The article emphasizes that, as of 2026-08-13, the public data still contained no complete independent third-party reproduction. The competing results in the table come from each model's publicly reported scores and cannot replace measurements under a unified harness.
Grok 4.6