Real-world results are in for GPT-6 Sol (Max) by @OpenAI. It landed #4 in the Code…

This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
Real-world results are in for GPT-6 Sol (Max) by @OpenAI. It landed #4 in the Code Arena: WebDev with 1689 pts, and has reshaped the Pareto frontier, at $8/M tokens (blended input/output)! This performance is on par with Claude Opus 5 (Max), while costing less than half as much, and trails the frontier leader GPT-6 Astra (Max) by two positions on the Pareto for one-fifth the cost. GPT-6 Sol (Max) is a significant improvement from GPT-5.6 Sol (xHigh) at +72 pts, ranked at #17. It also improved in all categories: - Simulations (#18 -> #3) - Content Creation (#14 -> #3) - Gaming (#13 -> #3) - C
Posted by Arena.ai (227.7k followers) 1 h ago · 194 likes · 21.6k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- Tested no-CoT abilities of GPT-5.6 Luna/Sol vs GPT-6 Luna/Sol - the GPT-6 series shows a… — @N8Programs
- Anthropic lança Claude Opus 5.5 com desempenho do Fable 5.1, mas incluído na assinatura… — @Tec_Mundo
- Grok 4.7 is insanely fast. It turned a simple prompt into this fun mini-game in just… — @cb_doge
- MiniMax tokenizer, M3.1? 👀 — @eliebakouch
- We asked Claude Opus 5.5, GPT-6 Sol, Sol Pro, GPT-6 Luna and Luna Pro for the single… — @eachlabs
- OpenAI GPT‑6 Sol and Luna are on the APEX leaderboards. — @mercor
- Gemini 3.8 Flash tops this chess benchmark — @HarshithLucky3
- GPT-6 Sol isn't just cheaper than GPT-5.6, it's also faster! — @AdamHoltererer
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 44.8k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-23 18:46 UTC. Full method.