Meet Bonsai-2-27B-1bit-CRACK-GGUF: a 27B parameter model squeezed into 1-bit ternary…
This is a AI post classified by Jev as Open & local models (a model release), kept by the AI Radar because it carries real work, not commentary.
Meet Bonsai-2-27B-1bit-CRACK-GGUF: a 27B parameter model squeezed into 1-bit ternary weights. Runs on llama.cpp with CUDA and Metal. On-device AI just got wild.
Posted by Hugging Models (68.3k followers) 1 h ago · 1 likes · 551 views · view the original post on X. Kept by the AI Radar as Open & local models.
More AI work like this
- Mimo V2.6 Flash TP2 Running with a 1.8M KV Pool. — @Tech2Wild
- Meet empero-ai/Qwen3.8-35B-A3B-Distill: a 35B MoE model distilled for reasoning and… — @HuggingModels
- What’s going on here lol what’s the point of these ai reply bots — @lu__jasper
- Meet decider-2b: a 2B parameter decision model that runs in ONE pass. No chains, no… — @HuggingModels
- Xiaomi’s MiMo V2.6 Pro looks like a contender for the best new model to run locally. — @plotarmordev
- Here's a model that shouldn't be overlooked! — @mr_r0b0t
- I let Kev play Dwarf Fortress — @TelepathicPug
- llama.cppで起動したMiMo-V2.6-Distill-Qwen-9Bにやらせた。ラーメン屋ベンチの結果です。動画をご査収ください。 — @studio_yebisu
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 40.7k posts from 5k X accounts over the last 14 days, 1.7k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 01:37 UTC. Full method.