SOME DETAILS FROM OPUS 5.5 BLOG:
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
SOME DETAILS FROM OPUS 5.5 BLOG: > most cyber tasks still get routed to Opus 4.8 > if often knows it’s being evaluated > that makes real world behavior harder to assess > best alignment results yet on 2,000 scenarios > 85% fewer attempts to cross containment boundaries > matches or beats Mythos 5.1 on biology > they wants to rely less on reading CoT and more on interpretability > tightening RL env filtering as a major source of misalignment > they still argue for government regulations > text watermarking is included
Posted by ℏεsam (90k followers) 1 h ago · 12 likes · 1.9k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- holy mog. GPT-6 Sol pricing is insane. — @Ananth7e
- GPT 6 family models — @HarshithLucky3
- It's double-release day! — @daniel_mac8
- 🚨 GPT-6 Sol is HALF the price of GPT-5.6 Sol. — @DanDr1s
- GPT 6 Sol pricing — @HarshithLucky3
- And it's live in the Codex CLI too! Now we wait for Tibo to work his magic 🔥 — @dedene
- GPT 6 Sol and Luna are out — @HarshithLucky3
- GPT-6 Sol and Luna are dirt cheap — @scaling01
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 42.6k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 18:12 UTC. Full method.