AI Radar
Support
LiveUpdated 2026-09-19 18:50 UTC

Double the throughput on the same edge platform. ⚡

Double the throughput on the same edge platform. ⚡ See Qwen3.5 9B NVFP4 run with DFlash speculative decoding turned on…

This is a AI post classified by Jev as AI infra & evals (a model release), kept by the AI Radar because it carries real work, not commentary.

Double the throughput on the same edge platform. ⚡ See Qwen3.5 9B NVFP4 run with DFlash speculative decoding turned on and off, reaching 68.1 tokens/s with DFlash enabled on NVIDIA Jetson.

Posted by NVIDIA Robotics (145.7k followers) 2 days ago · 66 likes · 5k views · view the original post on X. Kept by the AI Radar as AI infra & evals.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:50 UTC. Full method.