AI Radar
Support
LiveUpdated 2026-09-23 15:40 UTC

MiMo-V3 is getting a new architecture. The core of it, HySparse2, is out today.

MiMo-V3 is getting a new architecture. The core of it, HySparse2, is out today. Less prefill, a smaller KV cache,…

This is a AI post classified by Jev as AI research (a launch), kept by the AI Radar because it carries real work, not commentary.

MiMo-V3 is getting a new architecture. The core of it, HySparse2, is out today. Less prefill, a smaller KV cache, better long-context retrieval—and we got all three at once. Compared with MiMo-V2.6's Hybrid SWA architecture: • 5.02× lower prefill FLOPs at 1M tokens • 4.5× smaller KV cache at 1M tokens • Better MRCRv2 and RULER-v2 scores, plus lower AgentPPL and LongPPL Why build a new architecture? Agentic inference is a very different workload. Each round, a short action can return a long observation that needs to be prefilled, while the context keeps growing. That puts prefill cost, KV-ca

Posted by Fuli Luo (84.1k followers) 1 h ago · 1.3k likes · 47.7k views · view the original post on X. Kept by the AI Radar as AI research.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 44.5k posts from 5k X accounts over the last 14 days, 1.9k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-23 15:40 UTC. Full method.