AI Radar
Support
LiveUpdated 2026-09-22 19:32 UTC

Benchmarks GPT-6 Sol & Luna:

Benchmarks GPT-6 Sol & Luna: • FrontierCode: 48.4% vs. Fable 5.1’s 48.7%, both at xhigh. $1.37 vs. $9.27 per task,…Benchmarks GPT-6 Sol & Luna: • FrontierCode: 48.4% vs. Fable 5.1’s 48.7%, both at xhigh. $1.37 vs. $9.27 per task,…

This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.

Benchmarks GPT-6 Sol & Luna: • FrontierCode: 48.4% vs. Fable 5.1’s 48.7%, both at xhigh. $1.37 vs. $9.27 per task, roughly 85% cheaper. • DeepSWE: 68.8% at max vs. Fable 5’s best result of 69.9% at xhigh. $2.74 vs. $13.41 per task, roughly 80% cheaper. • AutomationBench: 33.2% at xhigh vs. Fable 5.1 with Opus 5 fallback at 31.4% on max. $0.27 vs. at least $2.45 per task. GPT-6 Luna: • DeepSWE: 66.6% at max vs. Fable 5’s 65.4% at medium. $0.22 vs. $6.09 per task, roughly 96% cheaper. API pricing per million input/output tokens: • Sol: $2 / $10 • Luna: $0.10 / $0.50

Posted by Chubby♨️ (145.6k followers) 1 h ago · 511 likes · 43k views · view the original post on X. Kept by the AI Radar as Frontier models.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 42.9k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 19:32 UTC. Full method.