ChessBench
ChessBench is For those of you who are into supporting independent benchmarks. It is ranked #215 on the AI Radar, in Frontier models, first seen 6 days ago and shared in 3 posts (1.1k views).
What ChessBench says about itself
Support ChessBench — fund API costs and project time so more models get benchmarked, faster. Sponsor monthly or make a one-time contribution.
Support | ChessBench Chess Bench [PREVIEW] A New Chess Benchmark for Language Models Support ChessBench Independently benchmarking a language model takes hundreds of games and, for newer frontier models, costs hundreds -- sometimes thousands -- of dollars in API fees. Your support helps test more models, faster and more thoroughly. GitHub Sponsors One-time or monthly support with tiers and a public sponsor badge. Become a sponsor on GitHub → Stripe Pay any amount, no account required. Make a one-time contribution , or give monthly: $5 · $10 · $20 · $50 Community Tell us which models or matchups you'd like to see, ask questions, share…
What people said about ChessBench on X
Gemini 3.8 Flash tops this chess benchmark with a huge margin is 3.8 flash, that good at chess....
— @HarshithLucky3, 50 min ago · 11 likes · see the post
Opus 5.5 is 4th on ChessBench Gemini does really well weirdly
— @AdamHoltererer, 1 h ago · 6 likes · see the post
Hey ChessBench, how much API cost would it be to add Astra? Maybe we can all pitch in
— @AdamHoltererer, 6 days ago · 2 likes · see the post
Alternatives to ChessBench
- Intelligence Benchmarking — Detailed intelligence benchmarking methodology for LLM quality evaluations.
- Qwen Studio — Qwen Studio offers comprehensive functionality spanning chatbot, image and video understanding, image generation,…
- Claude — Claude is Anthropic's AI, built for problem solvers. Tackle complex challenges, analyze data, write code, and think…
- 阶跃星辰 — Model page
- www-cdn.anthropic.com —
- Qwen3.8-Omni-Flash — Qwen’s next-generation native omni-modal model supports context lengths of up to 1M tokens and natively accepts text,…
ChessBench in numbers
- Rank on the AI Radar: #215 of 1874
- Shared in 3 posts by 2 accounts: @HarshithLucky3, @AdamHoltererer
- 1.1k views on those posts
- First seen 6 days ago, last shared 50 min ago
- Pricing seen by Jev: not stated
- Market: Frontier models
FAQ
What is ChessBench?
For those of you who are into supporting independent benchmarks It was first shared on X 6 days ago and is ranked #215 on the AI Radar.
Is ChessBench free?
Pricing is not stated on the page we read.
Who shared ChessBench?
2 accounts on X, including @HarshithLucky3, @AdamHoltererer, in 3 posts totalling 1.1k views.
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 44.8k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-23 18:46 UTC. Full method.