We compared the top model from each family by net improvement score and median cost per…
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
We compared the top model from each family by net improvement score and median cost per task. Family is defined as variants within the same model base name. GPT-6 Astra and Claude Fable 5.1 stand out: both cost far more than other models from their respective labs despite relatively close performance. By @OpenAI: - GPT-6 Astra (Max): +11.7% | $3.94/task - GPT-5.6 Sol (xHigh): +7.0% | $1.03/task By @AnthropicAI: - Claude Fable 5.1 (Max): +13.7% | $4.40/task - Claude Opus 5 (High): +10.2% | $2.07/task
Posted by Arena.ai (226.6k followers) 2 days ago · 363 likes · 33.4k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- Les premières pièces totalement modélisés par GPT6-Astra — @DFintelligence
- gpt 5.6 terra = gpt 6 astra — @notjazii
- wow — @HarshithLucky3
- 🚨 Gemini 4 Pro is absolutely cooking — @Mr_Salio
- anthropic is about to mog everyone with new models — @notjazii
- woke people still don’t want to accept AIs good potential, while AI be like to them🖕🏻 — @SciTechera
- @AdamHoltererer — @AdamHoltererer
- GPT 6 Astra Pro builds OCD simulator: — @AdamHoltererer
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:50 UTC. Full method.