A GLM-5.3 753B coding/cyber model that now fits on 2× Sparks
This is a AI post classified by Jev as Open & local models (a tool drop), kept by the AI Radar because it carries real work, not commentary.
A GLM-5.3 753B coding/cyber model that now fits on 2× Sparks or If you have 2 Sparks, an M5 Ultra, or 3× 6000s, you can run it locally. - REAP scores real contribution - Then EXL3 3-bit on the survivors. - KL vs full BF16: 0.511 - Top-1 next-token match: 78% - deleted 88 experts per layer with REAP. cut that actually loads, keep-168/ 256 experts, 3.0 bpw EXL3, and lands on hardware a serious lab can actually own. - http://huggingface.co/0xSero/GLM-5.3-500B-EXL3-3.0bpw
Posted by Md Ismail Šojal 🕷️ (56k followers) 2 h ago · 0 likes · 447 views · view the original post on X. Kept by the AI Radar as Open & local models. Tools mentioned: glm-5.3-500b-exl3-3.0bpw.
More AI work like this
- New on the hub: MiMo-V2.6-Distill-Qwen-9B. A 9B image-text-to-text model built for… — @HuggingModels
- Paper thread on MiMo-V2.6, the newest open model and the latest to stream their RL run :) — @nrehiew_
- Ant Group’s Ling 3.0 Flash Fin is a finance-specialized open-weight model that delivers… — @ValsAI
- DeepSeekHarnessにサブスクのCODEXを設定できた — @Tono_Ken3
- Koi pond local battle. Same 2 DGX Sparks each. Three Flash models. — @WescheNex1q
- Meet Qwen-72B: a massive 72B parameter text generation model. It handles both Chinese… — @HuggingModels
- I had Opus 5.5 build a tiny Sims town where every resident's dialogue is written live by… — @Tech2Wild
- Meet Qwen3.8-27B iMatrix NVFP4 MTP GGUF: a multimodal beast that reads images AND text.… — @HuggingModels
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 43.5k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 23:36 UTC. Full method.