Meet Confucius4-R2T2, a streaming ASR model that turns speech into text in real time.…
This is a AI post classified by Jev as Voice AI (a model release), kept by the AI Radar because it carries real work, not commentary.
Meet Confucius4-R2T2, a streaming ASR model that turns speech into text in real time. Built on Qwen3 ASR, it's designed for low latency, so your apps can transcribe live audio without the wait. This is a big deal for voice interfaces.
Posted by Hugging Models (68.2k followers) 20 h ago · 13 likes · 1.9k views · view the original post on X. Kept by the AI Radar as Voice AI.
More AI work like this
- SpaceXAI just released Grok Voice Transcribe 2.0.....and it already ranks #1 for BOTH… — @XFreeze
- 开源社区又在整大活: — @mylifcc
- なんてこった!! — @studio_yebisu
- This is the best open source alternative to ElevenLabs and WisprFlow that you can runs… — @hasantoxr
- Meet Qwen3.8-LiveTranslate, Qwen's next-generation real-time simultaneous interpretation… — @Alibaba_Qwen
- Grok Bot Voice is now rolled out to 100% of mobile users (both iOS and Android….i just… — @XFreeze
- Grok Voice Transcribe 2.0 is now #2 on Voice Code Bench—up 13pp and 12 spots from its… — @ValsAI
- https://openai.com/index/continuous-voice-interaction-with-gpt-live/ — @TheRealAdamG
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:08 UTC. Full method.