AI Radar
Support
LiveUpdated 2026-09-24 02:41 UTC

I evaluated the latest models, GPT-6 Sol, Luna, and Claude Opus 5.5 on induction (post…

I evaluated the latest models, GPT-6 Sol, Luna, and Claude Opus 5.5 on induction (post from yesterday). GPT-6 Sol is…

This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.

I evaluated the latest models, GPT-6 Sol, Luna, and Claude Opus 5.5 on induction (post from yesterday). GPT-6 Sol is doing very well, somewhat behind Opus 5.5 but at 1/2 the price. GPT-6 Luna is doing slightly better than GPT-5.6 Luna. 14 responses are still running on high thinking effort after failing on max and xhigh. I don't expect them to significantly raise the rate of success. Overall, I think this is a success for the new GPT-6 models. The new Sol is significantly stronger than the old Sol at 2/3 the price, and slightly behind Opus 5.5. Luna is slightly above the previous Luna at 1/3

Posted by Serafim Batzoglou (3.7k followers) 1 h ago · 9 likes · 525 views · view the original post on X. Kept by the AI Radar as Frontier models. Tools mentioned: concept-synth.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 45.7k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-24 02:41 UTC. Full method.