We evaluated @claudeai Opus 5.5 on Box's Complex Work Eval. The standout finding is…
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
We evaluated @claudeai Opus 5.5 on Box's Complex Work Eval. The standout finding is efficiency. → 30% faster than Opus 5 → A third of the tokens → Answers 42% shorter with the findings intact What disappeared was restatement and preamble. Opus 5.5 simply led with the conclusions. Accuracy held across Consumer Products, Financial Services, and Technology. Claude Opus 5.5 is coming to Box AI soon.
Posted by Box (80k followers) 4 h ago · 21 likes · 253.7k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- if you’re using opus 5.5, steal this prompt... — @Av1dlive
- Claude Opus 5.5 system card: — @rohanpaul_ai
- anthropic won big today — @notjazii
- One launch deal for 5 new models. — @RespanAI
- Claude Opus 5.5 on Higgsfield turned one house photo and its floor plans into a full 3D… — @higgsfield_ai
- A quick recap of what we saw today, and an answer to the question of whether there was a… — @kimmonismus
- Grok 4.7 (xHigh) by @SpaceXAI just landed at #10 in Code Arena: WebDev with 1632 pts. — @arena
- Claude Opus 5.5 vs GPT-6 Sol in 3D game development. — @higgsfield_ai
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 43.1k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 20:50 UTC. Full method.