GPT-6 Sol review navigator
Official benchmarks, independent analysis, and community reports about GPT-6 Sol, clearly separated from Tabbit's own testing.
Official
1 source-checked resourcesMedia
3 source-checked resourcesGPT-6 Sol: Artificial Analysis on Cost Efficiency and Hallucination Measurement
Artificial Analysis's published results show that GPT-6 Sol max scores 48 on the Intelligence Index and 57 on the Coding Agent Index. The former costs about half as much per task as GPT-5.6 Sol max, while the Coding Agent Index score is 2 points higher. The AA.
GPT-6 Sol on AI IQ: Model Profile and Benchmark Coverage
The AI IQ model page gives GPT-6 Sol an estimated overall IQ of 136, but only academic reasoning, coding reasoning, and reliability have direct benchmark results among the six dimensions. The remaining dimensions and some missing benchmarks are estimated or im.
GPT-6 Sol: Artificial Analysis Comparison Across Six Configurations
Artificial Analysis's release page lists six GPT-6 Sol configurations. The Intelligence Index v4.3.2 ranges from 28 for Non-reasoning to 48 for max; output speed ranges from 109 to 129 tokens/s; and the weighted average cost per Intelligence Index task ranges .
Community
3 source-checked resourcesGPT-6 Sol: KillSwitch-Bench Adversarial Esoteric-Language Coding Agent Benchmark
The KillSwitch-Bench leaderboard lists GPT-6 Sol (Codex harness) with a composite score of 28.0%, a cost of $0.42 per task, and a runtime of 2m15s. The 28.0% is a composite score based on pass rate, cost, and code size; it must not be read as accuracy or the t.
GPT-6 Sol: Arena Comparison with GPT-5.6 Sol
Arena says Peter Gostev directly compared GPT-6 Sol with GPT-5.6 Sol under matched conditions: both used the same prompts and were set to the maximum reasoning level. The post says the evaluation covers generated results, total token counts, and wall-clock tim.
Reddit Single-Task Report: GPT-6 Sol xhigh vs. an Older Sol Model
The poster says GPT-6 Sol xhigh made its first mistake on an unspecified existing task, while GPT-5 through GPT-5.5 series models had not made one before. This is a single personal report that cannot be independently verified, and does not establish the overal.
GPT-6 Sol
Use and compare models in Tabbit
Official benchmarks, independent analysis, and community reports about GPT-6 Sol, clearly separated from Tabbit's own testing.