Ran FlappyBench on Qwen3.8-Omni-Flash, DeepSeek-V4.1-Flash, and Gemini 3.8 Flash with…
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
Ran FlappyBench on Qwen3.8-Omni-Flash, DeepSeek-V4.1-Flash, and Gemini 3.8 Flash with the same design prompt. 🔹 Qwen 3.8 Omni Flash: 9/10 · $0.013 · smooth gameplay and cheaper 🔹 DS V4.1 Flash: 9/10 · $0.0089 one-shot with the lowest cost 🔹 Gemini 3.8 Flash: 5/10 · $0.205 · 4-5 iterations but still hard-coded gameplay and costs 23x more
Posted by Command Code (23.7k followers) 1 h ago · 61 likes · 4.1k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- Claude Fable 5.2 @claudeai @ClaudeDevs you guys seriously cooked with this model — @chetaslua
- Chat GPT-6 Astra solve 108 year old unsolved WWI german code for first time in history. — @Vvikramai
- The most interesting thing about this release, to me, is the implication that there had… — @teortaxesTex
- We’ve been quiet for too long. It’s time for a change. — @sheriyuo
- Introducing Step 5 Preview: Advancing the Pareto Frontier. — @StepFun_ai
- @Luna11054 — @Luna11054
- - You wake up — @0x0SojalSec
- in 1 video Understand Jev model better than most of peoples. — @0x0SojalSec
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 31.7k posts from 4.8k X accounts over the last 14 days, 1.3k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-20 04:05 UTC. Full method.