So I tested Pareto by Unbiased (ex-Union Alpha) on a real work: 2 repos, 105 planted…
This is a AI post classified by Jev as Frontier models (a model release), kept by the AI Radar because it carries real work, not commentary.
So I tested Pareto by Unbiased (ex-Union Alpha) on a real work: 2 repos, 105 planted bugs. Find and fix what you can. Score vs cost: GPT-6 Astra (max): 45, $33.03 Fable 5.1 (max): 43, $77.55 Muse Spark 1.3 (max): 32.2, $18.11 Pareto by Unbiased 30.7, $4.81 ⭐ Grok 4.6 (xhigh): 28.7, $18.60 Opus 5 (max): 27, $51.33 Fast and strong. Turns out it's a composite model.
Posted by Paweł Huryn (130.9k followers) 1 days ago · 119 likes · 17.2k views · view the original post on X. Kept by the AI Radar as Frontier models.
More AI work like this
- TikTok reacts to this: — @_AashishReddy
- Lyria 3.5 from Google DeepMind is now live on fal — @fal
- Tested Jev vs GPT-5.6 for predicting my taste in books. — @venturetwins
- Wife wanted 5-yo involved, so Fable gave me a bunch of reading questions to give him to… — @kevinnbass
- opus 5.2 learned to write poems — @aliceisplaying
- meanwhile grok first try: — @thekitze
- Les premières pièces totalement modélisés par GPT6-Astra — @DFintelligence
- gpt 5.6 terra = gpt 6 astra — @notjazii
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 31.3k posts from 4.7k X accounts over the last 14 days, 1.3k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 21:10 UTC. Full method.