AI Radar
Support
LiveUpdated 2026-09-20 02:50 UTC

Full Kimi K3 (2.8T) frontier model running locally on 16× NVIDIA GB10 Spark Desk…

Full Kimi K3 (2.8T) frontier model running locally on 16× NVIDIA GB10 Spark Desk cluster, Not a data center. - 30 t/s…Full Kimi K3 (2.8T) frontier model running locally on 16× NVIDIA GB10 Spark Desk cluster, Not a data center. - 30 t/s…

This is a AI post classified by Jev as AI infra & evals (a model release), kept by the AI Radar because it carries real work, not commentary.

Full Kimi K3 (2.8T) frontier model running locally on 16× NVIDIA GB10 Spark Desk cluster, Not a data center. - 30 t/s single stream - 87 t/s aggregate at C8 - 136 t/s peak claimed - Prose llama-bench holding 20–23 t/s generate even at long context. You don’t need a hyperscaler rack to run a 2.8T model anymore.

Posted by Md Ismail Šojal 🕷️ (55.6k followers) 1 h ago · 4 likes · 837 views · view the original post on X. Kept by the AI Radar as AI infra & evals.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 31.6k posts from 4.8k X accounts over the last 14 days, 1.3k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-20 02:50 UTC. Full method.