Grok-4.7 first test results are in and it's not looking good.
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
Grok-4.7 first test results are in and it's not looking good. Floating island with pagoda in three.js. Bad compared to DeepSeek-v4.1-flash result i got. What makes this even worse is that, this is not one-shot, this was done from Grok build, a complete harness, few turns and with tool calls. Overall quality is off, appears like a very lazy and approximate implementation.
Posted by AJ (7.7k followers) 3 h ago · 16 likes · 764 views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- Damn. I’m late to the party. Plane just arrived in SF: Grok 4.7 looks super amazing! — @kimmonismus
- just taking a moment to appreciate how far these 9B models have come. — @suchenzang
- BREAKING: MiMo-V2.6-Pro by @XiaomiMiMo lands at #8 overall (#3 open-weight) on Design… — @DesignArena
- 小米 MiMo-V2.6 — @dongxi_nlp
- Chinese models are very close to saturating CritPt (it's broken and in reality 30% is… — @teortaxesTex
- Grok 4.7 is now available in Devin. — @cognition
- omg xiaomi cooked hard with mimo 2.6 — @notjazii
- The difference between the frontier of closed source and open source has been reduced by… — @himanshustwts
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 40.3k posts from 5k X accounts over the last 14 days, 1.7k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-21 21:52 UTC. Full method.