LiveUpdated 2026-09-24 11:00 UTC
V4.1,GLM-5.3,Step-5の推論カーネルの記憶を使い
This is a AI post classified by Jev as AI infra & evals (a model release), kept by the AI Radar because it carries real work, not commentary.
Mimo-V2.6-Flash(揉光-無検閲) V4.1,GLM-5.3,Step-5の推論カーネルの記憶を使い RTXPro4000+CPU推論で23TPSまで出てきた 目標は30TPSだが果たして
Posted by となりのトトノ🏯Local LLM | Tonoken3 (5.2k followers) 1 h ago · 0 likes · 1 views · view the original post on X. Kept by the AI Radar as AI infra & evals.
More AI work like this
- Training your own Jev in minutes. — @DataChaz
- The right model for the right AI task. — @alibaba_cloud
- Nice paper showing how to re-evaluate a production agent at a fraction of the cost. — @omarsar0
- I made a Jev competitor called Jev-Huyev: https://www.chapterpal.com/jev-huyev — @burkov
- Ok, I made a Jev competitor called Jev-Huyev: https://www.chapterpal.com/jev-huyev — @burkov
- Note that Jev-like classifier doesn't imply Jev-like capability - on a wide variety of… — @N8Programs
- This is basically what you should be seeing on GLM-5.3-Flash with 4x DGX Sparks. If… — @mmastrac
- Pushing higher GLM 5.3 Flash TP4 DGX sparks — @TechMDAI
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 46k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-24 11:00 UTC. Full method.