update on the same bundle. a new prefill kernel landed in the prismml fork this morning,…
This is a AI post classified by Jev as Open & local models (a tool drop), kept by the AI Radar because it carries real work, not commentary.
update on the same bundle. a new prefill kernel landed in the prismml fork this morning, i rebased onto it and rebuilt, and prompt processing went from 267 to 522 tok/s on the 3060. this one is purely how fast it reads your prompt. a 32k prompt that took two minutes to load now takes one. same weights, same gguf, nothing to re-download except the 492 mb bundle. i ran both builds back to back in one session rather than comparing against yesterday's number.
Posted by Sudo su (36.4k followers) 1 h ago · 4 likes · 1.3k views · view the original post on X. Kept by the AI Radar as Open & local models.
More AI work like this
- Xiaomi have distilled MiMo V2.6 into Qwen3.5-9B and have introduced an MTP head into it… — @Biggest
- Introducing Hotdog-27B, a specialized bipartite classifier. — @michael_chomsky
- We got new image gen model releases from @AntLingAGI — @ItsmeAjayKV
- Meet a 100M parameter transformer trained on the BabyLM dataset. It is a text generation… — @HuggingModels
- Just launched my usual Qwen 3.8 27B configs and, for a brief moment, couldn't understand… — @Anbeeld
- We’re open-sourcing the Ming-Image-0.1-Design family: — @AntLingAGI
- Introducing https://theopenfrontier.com! — @nutlope
- Meet Lanni-ni/hard_3gram_4_6_384_babylm_10m_seed44: a text-generation transformer… — @HuggingModels
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 43k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 20:11 UTC. Full method.