AI Radar
Support
LiveUpdated 2026-09-19 18:08 UTC

Meet Bonsai 2 27B: a 27B model squeezed into 2-bit ternary with GGUF for llama.cpp. Runs…

Meet Bonsai 2 27B: a 27B model squeezed into 2-bit ternary with GGUF for llama.cpp. Runs on CUDA, Metal, and…

This is a AI post classified by Jev as Open & local models (a model release), kept by the AI Radar because it carries real work, not commentary.

Meet Bonsai 2 27B: a 27B model squeezed into 2-bit ternary with GGUF for llama.cpp. Runs on CUDA, Metal, and on-device. Hybrid attention + PrismML. Big brain, tiny footprint, local speed. 1,434 downloads and counting.

Posted by Hugging Models (68.2k followers) 21 h ago · 59 likes · 3.8k views · view the original post on X. Kept by the AI Radar as Open & local models.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:08 UTC. Full method.