deepseek-v4.1-flash-vllm-dgx-spark
deepseek-v4.1-flash-vllm-dgx-spark is DeepSeek-V4.1-Flash (552B MoE, MXFP4 experts, 1M ctx) on four NVIDIA DGX Sparks with vLLM TP4: Engram-on-disk patch, sm121 kernel build, launchers, measured numbers - tonyd2wild/DeepSeek-V4.1-Flash.... It is ranked #624 on the Dev Radar, in Hosting & infra, first seen 9 days ago and shared in 2 posts (3.7k views).
GitHub - tonyd2wild/DeepSeek-V4.1-Flash-vLLM-DGX-Spark: DeepSeek-V4.1-Flash (552B MoE, MXFP4 experts, 1M ctx) on four NVIDIA DGX Sparks with vLLM TP4: Engram-on-disk patch, sm121 kernel build, launchers, measured numbers · GitHub Skip to content Navigation Menu Sign in Appearance settings Search / Sign in Sign up Appearance settings You signed in with another tab or window. Reload to refresh your session. You signed out in another tab or window. Reload to refresh your session. You switched accounts on another tab or window. Reload to refresh your session. Dismiss alert {{ message }} tonyd2wild / DeepSeek-V4.1-Flash-vLLM-DGX-Spark Public…
What people said about deepseek-v4.1-flash-vllm-dgx-spark on X
I always seem to have more success with @Tech2Wild recipes. Reliable and well put together and always very quick support. Give him a follow if you haven’t already for your DGX sparks and 3090 goodies. 🔥
— @Blackwellboy, 8 days ago · 27 likes · see the post
Deepseek-v4.1-flash up and running, benchmark and visuals comparing to dsv4-flash-0731 coming up next
— @WescheNex1q, 9 days ago · 15 likes · see the post
Alternatives to deepseek-v4.1-flash-vllm-dgx-spark
- boat by ASCII — ascii box is now boat ( )
- Railway — Bring context from your work and personal accounts into the same conversation. Connect your accounts in the plugin…
- qwen3.8-flash-next-single-dgx-spark — Qwen3.8-Flash-Next on ONE DGX Spark (TP=1). Contribute to MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark development by…
- Nebius Token Factory — Start building
- qwen3.8-flash-next-exl3-dgx-spark-recipe — Serve Qwen3.8-Flash-Next (turboderp ExLlamaV3 pack) on one NVIDIA DGX Spark with vLLM + vllm-exl3 -…
- Qwen-Omni — Qwen-Omni,Alibaba Cloud Model Studio:Use Qwen-Omni over HTTP to understand text, images, audio, and video. Choose…
deepseek-v4.1-flash-vllm-dgx-spark in numbers
- Rank on the Dev Radar: #624 of 1578
- Shared in 2 posts by 2 accounts: @Blackwellboy, @WescheNex1q
- 3.7k views on those posts
- First seen 9 days ago, last shared 8 days ago
- Pricing seen by Jev: open source
- Market: Hosting & infra
FAQ
What is deepseek-v4.1-flash-vllm-dgx-spark?
DeepSeek-V4.1-Flash (552B MoE, MXFP4 experts, 1M ctx) on four NVIDIA DGX Sparks with vLLM TP4: Engram-on-disk patch, sm121 kernel build, launchers, measured numbers - tonyd2wild/DeepSeek-V4.1-Flash... It was first shared on X 9 days ago and is ranked #624 on the Dev Radar.
Is deepseek-v4.1-flash-vllm-dgx-spark free?
It is open source.
Who shared deepseek-v4.1-flash-vllm-dgx-spark?
2 accounts on X, including @Blackwellboy, @WescheNex1q, in 2 posts totalling 3.7k views.
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 14.4k posts from 4.7k X accounts over the last 21 days, 1.6k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 21:10 UTC. Full method.