智谱 Infra 是真的强啊,刚刚又推出了 GLM-5.3-FlashX,是此前 GLM-5.3-Flash 的高速版,官宣最高速度 200 tokens/s。
This is a AI post classified by Jev as AI infra & evals (a model release), kept by the AI Radar because it carries real work, not commentary.
智谱 Infra 是真的强啊,刚刚又推出了 GLM-5.3-FlashX,是此前 GLM-5.3-Flash 的高速版,官宣最高速度 200 tokens/s。 💰价格约为 Flash 的 2.5 倍:输入 $0.37 / 百万 Tokens,输出 $1.25 / 百万 Tokens,缓存命中 $0.075 /百万 Tokens。
Posted by FeiZ (8.1k followers) 1 days ago · 121 likes · 31.5k views · view the original post on X. Kept by the AI Radar as AI infra & evals.
More AI work like this
- Trains your own large language model from scratch using plain PyTorch. — @tom_doerr
- Here's a useful use-case for @typesafeai for a change. — @iam_zachi
- This is genius. — @daniel_mac8
- Which GPU should you use for embedding workloads? — @runpod
- Big update to the @Alibaba_Qwen Qwen3.8-Flash-Next single DGX Spark recipe! — @jvr0x
- 🔰 The Databricks Advanced Learning Festival runs September 16 through October 14, with… — @databricks
- Q: What is a trace? — @HamelHusain
- Inference scaling part 1. — @rasbt
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:45 UTC. Full method.