AI Radar
Support
LiveUpdated 2026-09-19 18:50 UTC

Today’s LLMs still write like typewriters: one token at a time. This sequential process…

Today’s LLMs still write like typewriters: one token at a time. This sequential process creates a hard inference…

This is a AI post classified by Jev as AI research (a model release), kept by the AI Radar because it carries real work, not commentary.

Today’s LLMs still write like typewriters: one token at a time. This sequential process creates a hard inference bottleneck. We're introducing Uno, a diffusion-augmented LLM that delivers autoregressive quality at diffusion speed. It’s a lossless speedup method that accelerates generation without degrading response quality. With Uno, K2-Horizon-7B outperforms state-of-the-art diffusion methods in both quality and throughput, delivering up to a 2.2× speedup with no loss in quality. Paper: https://arxiv.org/abs/2609.04010 Model available at: https://huggingface.co/IFM/K2-Horizon-7B-Uno

Posted by Institute of Foundation Models (3.5k followers) 2 days ago · 1k likes · 220.2k views · view the original post on X. Kept by the AI Radar as AI research. Tools mentioned: k2-horizon-7b-uno.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:50 UTC. Full method.