Google's index shows that the Rails team compared 8 models on 21 atomic tasks. The summary reports these results for Luna: 73% of runs succeeded with the default medium reasoning effort, the total cost of 63 runs was approximately 90 cents, and it was labeled the cheapest model.
Agents on Rails: We ran 8 models against 21 atomic tasks to see which were ... Cheapest: @OpenAI GPT-5.6 Luna. 73% of ru… This is a necessary excerpt; read the original source for full context.
The body of the X post did not expand in the Tabbit international edition, so this note preserves only the original excerpt visible in Google's index. The task list, success criteria, costs for the other models, and runtime environment are missing; therefore, the 73% figure must not be treated as a general coding pass rate.
GPT-5.6 Luna