NetEase-Youdao releases Confucius4-R2T2, a true-streaming ASR model for live captions,…
This is a AI post classified by Jev as Voice AI (a model release), kept by the AI Radar because it carries real work, not commentary.
NetEase-Youdao releases Confucius4-R2T2, a true-streaming ASR model for live captions, simultaneous translation, and voice agents. 🤖 https://modelscope.ai/models/netease-youdao/Confucius4-R2T2 🏆 R2T2 achieves SOTA latency and recognition quality among the evaluated open-source models, while remaining competitive with leading closed-source systems. ⚡ Configurable 80ms–2s chunks deliver 200–600ms average latency with near-offline recognition accuracy. 📝 Append-only decoding commits stable text without revising earlier words, avoiding transcript flicker and giving downstream agents reliable inpu
Posted by ModelScope (15.6k followers) 1 h ago · 17 likes · 1.1k views · view the original post on X. Kept by the AI Radar as Voice AI. Tools mentioned: Ling-3.0-flash-Fin.
More AI work like this
- NVIDIA research papers are on fire recently! — @omarsar0
- New model: qwen-audio-3.1-realtime-plus — @aitrackerbot
- Meet Qwen3.8-LiveTranslate, Qwen's next-generation real-time simultaneous interpretation… — @alibaba_cloud
- SpaceXAI just released Grok Voice Transcribe 2.0.....and it already ranks #1 for BOTH… — @XFreeze
- 开源社区又在整大活: — @mylifcc
- なんてこった!! — @studio_yebisu
- This is the best open source alternative to ElevenLabs and WisprFlow that you can runs… — @hasantoxr
- Meet Qwen3.8-LiveTranslate, Qwen's next-generation real-time simultaneous interpretation… — @Alibaba_Qwen
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 34.2k posts from 4.8k X accounts over the last 14 days, 1.4k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-20 10:41 UTC. Full method.