Introducing GPT-6 Sol and Luna, bringing the advances behind Astra to faster, more…
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
Introducing GPT-6 Sol and Luna, bringing the advances behind Astra to faster, more affordable models. ✨ 💻 Stronger coding and computer use 🎯 Improved factuality and alignment 💬 Clearer answers with less jargon > On AutomationBench, Sol at xhigh effort outperforms Claude Opus 5 at max effort at just 9% of its cost per task. > On OSWorld 2.0 offline, Luna at max effort exceeds GPT-5.6 Sol at medium effort at one-tenth the cost. > Sol makes about half as many mistakes as GPT-5.6 Sol on our internal factuality evaluation of conversations where users previously flagged errors. > Improved promp
Posted by Vaibhav (VB) Srivastav (59.6k followers) 1 h ago · 561 likes · 56.3k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- .@AnthropicAI just shipped Claude Opus 5.5! — @merge_api
- Well, Opus 5.5 failed this test — @literallydenis
- 给大家带来claude opus 5.5 的前端测试结果. 直接说结论, 我怀疑现在opus 5.5就是 fable 5.1 的量化版或者自家蒸馏版, 可以直接看我的视频,… — @karminski3
- Anthropic Launches Claude Opus 5.5 With Fable-Level Performance at a Lower Price… — @MacRumors
- This week’s model drop just rewrote the leaderboard. — @StatsWire
- Use GPT-6 Sol and Luna with AI SDK — @aisdk
- AA benchmark has to be a joke at this point — @haider1
- NO MORE EM-DASHES — @vikktorrrre
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 42.9k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 19:32 UTC. Full method.