Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Reviews and evidence

GPT-5.6 Luna · Official source · Vendor report

GPT-5.6 Sol, Terra, and Luna: Three-Tier Reddit Benchmarks and Routing Recommendations

This evidence note covers “GPT-5.6 Sol, Terra, and Luna: Three-Tier Reddit Benchmarks and Routing Recommendations” under stated conditions; its version, sample, and runtime limits do not support a universal ranking or current production guarantee.

Unverified: the original source could not be rechecked. Historical figures below are not current verified results.

Official sourceVendor reportEdited 2026-09-20

Test conditions

Model and version
GPT-5.6 Luna; do not merge with other versions, reasoning tiers, or harnesses.
Provider / environment
Reddit, r/OpenAI; the original conditions do not establish one controlled retest.
Collection boundary
The source note was collected on 2026-08-17/18; the original page was not reopened this round, so dynamic facts remain unverified.

Key data and applicable tasks

Summary

The post compiles pricing and multiple benchmarks for the three model tiers and offers routing recommendations: use Terra as the default model for most workloads, reserve Sol for the hardest agentic/terminal tasks, and use Luna for high-frequency pipelines. The post also sparked discussion about hallucinations, subscription quotas, and whether benchmarks represent the real-world experience.

Results

  • Original prices in the post: Sol $5/$30, Terra $2.50/$15, and Luna $1/$6 (per million input/output tokens; prices have since changed and should not be treated as current pricing).

  • Terminal-Bench 2.1: Sol 88.8%, Terra 87.4%, and Luna 84.7%.

  • AA Coding Agent Index: Sol 80, Terra 77.4, and Luna 74.6.

  • DeepSWE value: the post says Luna delivers approximately 24 points per $1, compared with approximately 4.5 points for Opus 4.8.

  • Commenters noted that Sol may hallucinate in actual use, while others felt that Luna Max was close to Sol Medium, although the speed and efficiency are not exactly the same.

Original article

OpenAI shipped GPT-5.6 to GA on July 9 — three tiers (Sol, Terra, Luna) that can evolve on independent cadences, plus ne… This is a necessary excerpt; read the original source for full context.

What this supports

  • The three-tier post suggests Terra by default, Sol for hard planning, and Luna for high-frequency work, using historical $1/$6 prices as context.

What this does not support

  • Historical prices and subjective tiers are not current official pricing, a unified ranking, or a stable routing guarantee; the full method is not public.

Method, limits, and reproduction

The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.

For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.

Original source

Reddit, r/OpenAI · u/docdavkitty; the linked article is credited to the-agent-report.com · Original publication date Unknown · Site edit date 2026-09-20

Open original source

GPT-5.6 Luna

Compare GPT-5.6 Luna in Tabbit

Download the Tabbit client to check model access

Related reviews

GPT-5.6 Luna Benchmarks & Pricing (Public Benchmarks & Pricing)This evidence note covers “GPT-5.6 Luna Benchmarks & Pricing (Public Benchmarks & Pricing)” under stated conditions; its version, sample, and runtime limits do not support a universal ranking or current production guarantee.I Benchmarked GPT-5.6 Sol/Luna/Terra by Role: Role-Based EvaluationThis evidence note covers “I Benchmarked GPT-5.6 Sol/Luna/Terra by Role: Role-Based Evaluation” under stated conditions; its version, sample, and runtime limits do not support a universal ranking or current production guarantee.GPT-5.6 Luna Semgrep IDOR Security Benchmark and Cost per True PositiveThis evidence note covers “GPT-5.6 Luna Semgrep IDOR Security Benchmark and Cost per True Positive” under stated conditions; its version, sample, and runtime limits do not support a universal ranking or current production guarantee.GPT-5.6 Luna and Gemini 3.6 Flash: A Cost Counterexample in Document-Vision TasksThis evidence note covers “GPT-5.6 Luna and Gemini 3.6 Flash: A Cost Counterexample in Document-Vision Tasks” under stated conditions; its version, sample, and runtime limits do not support a universal ranking or current production guarantee.Get Started with OpenAI GPT-5.6 on Amazon Bedrock: Reasoning, Tool Calling, and CachingAWS's article positions Luna as a high-throughput, low-latency model for classification, summarization, and routing, and uses Responses API examples to show how to set reasoning effort, call tools, carry the model's output in full into the next turn, and cache。The Builder's Guide to GPT-5.6: Luna's Model Selection, Agent Orchestration, and CachingTurn “The Builder's Guide to GPT-5.6: Luna's Model Selection, Agent Orchestration, and Caching” into a bounded task entry with explicit inputs, runtime context, output format, and acceptance checks; confirm the model version and source limits before use.GPT-5.6 Prompting Guide: Luna's Work Contract and Model RoutingTurn “GPT-5.6 Prompting Guide: Luna's Work Contract and Model Routing” into a bounded task entry with explicit inputs, runtime context, output format, and acceptance checks; confirm the model version and source limits before use.Apply Occam’s Razor: Reducing Overengineering in Luna/Codex PromptsTurn “Apply Occam’s Razor: Reducing Overengineering in Luna/Codex Prompts” into a bounded task entry with explicit inputs, runtime context, output format, and acceptance checks; confirm the model version and source limits before use.