LiveUpdated 2026-09-24 01:35 UTC
The token usage chart here is really striking
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
The token usage chart here is really striking
Posted by Charlie Marsh (50.9k followers) 1 h ago · 29 likes · 2.7k views · view the original post on X. Kept by the AI Radar as Frontier models. Tools mentioned: Humanity's Last Exam.
More AI work like this
- There are some new top models leading our code review benchmark. — @Macroscope
- Claude Opus 5.5 is the greatest AI model ever released — @AlexFinn
- GLM-5.3 beat Space Bunny Alpha (the new stealth model launched today) on Newton’s cradle… — @rohanpaul_ai
- Anthropic’s guide to using Opus 5.5 has one central rule: stop asking it to think. Eddie… — @davidarngar
- Claude Opus 5.5 recreated an athlete’s long jump in 4D. — @higgsfield_ai
- Claude Opus 5.5 painted the past and future of our world in JavaScript + Higgsfield. — @higgsfield_ai
- Space Bunny performs TERRIBLE in physics 💀 — @aimlapi
- Okay, my first complete benchmark of Jev with my new Decision v1 eval suite on… — @morganlinton
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 45.7k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-24 01:35 UTC. Full method.