AI Radar
Support
LiveUpdated 2026-09-22 18:12 UTC

Artificial Analysis 的最新测试里,Opus 5.5 Max 单个任务平均消耗 11.9 万个输出 Token,其中光 reasoning 就用了 8.4 万…

🔥Claude Opus 5.5 有点逆天了。 Artificial Analysis 的最新测试里,Opus 5.5 Max 单个任务平均消耗 11.9 万个输出 Token,其中光 reasoning 就用了 8.4 万…

This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.

🔥Claude Opus 5.5 有点逆天了。 Artificial Analysis 的最新测试里,Opus 5.5 Max 单个任务平均消耗 11.9 万个输出 Token,其中光 reasoning 就用了 8.4 万 Token。 而GPT-6 Astra Max 每个 Intelligence Index 任务平均只消耗约 2.7 万 output tokens。 这接近 Astra 的 5 倍。 两家的路线差异现在已经非常明显了: Anthropic 在疯狂堆 test-time compute,OpenAI 则在疯狂压缩完成同类任务需要的推理长度。 这是目前榜单里最高的,甚至超过了以“巨能想”著称的 Qwen3.8 Max。 以前大家比谁更聪明。 接下来可能还得比达到同样的智能水平,到底要烧多少 Token。 and Test-time compute 还远远没撞墙。

Posted by Max For AI (39.9k followers) 1 h ago · 34 likes · 4.7k views · view the original post on X. Kept by the AI Radar as Frontier models.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 42.6k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 18:12 UTC. Full method.