Dev Radar
Support
LiveUpdated 2026-09-21 06:12 UTC

Big update to the @Alibaba_Qwen Qwen3.8-Flash-Next single DGX Spark recipe!

Big update to the @Alibaba_Qwen Qwen3.8-Flash-Next single DGX Spark recipe! 𝗪𝗵𝗮𝘁'𝘀 𝗻𝗲𝘄 🎁 • Measured on one…

This is a dev post classified by Jev as Hosting & infra (a tool drop), kept by the Dev Radar because it carries real work, not commentary.

Big update to the @Alibaba_Qwen Qwen3.8-Flash-Next single DGX Spark recipe! 𝗪𝗵𝗮𝘁'𝘀 𝗻𝗲𝘄 🎁 • Measured on one DGX Spark, 262K context, MTP k=3, aggregate tok/s at 1 / 2 / 4 / 8 streams: Prose: 38.0 / 61.1 / 89.2 / 117.4 Code: 53.8 / 87.5 / 131.8 / 180.2 • Long context holds: MTP keeps working at a 185K-token prompt (35.4 tok/s decode), prefill ~2,000 tok/s from 4K to 185K • ~1M-token KV pool at the full 262K context (FP8 KV) • NVIDIA's official NVFP4 checkpoint now runs on one Spark, with chat, tool calls and vision working • 24/7 mode: a supervisor restarts the server on a crash, detects

Posted by Javier • priv/acc (4.4k followers) 1 days ago · 115 likes · 69.2k views · view the original post on X. Kept by the Dev Radar as Hosting & infra.

More dev work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 18.9k posts from 4.9k X accounts over the last 21 days, 2.1k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-21 06:12 UTC. Full method.