Model Releases
What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • Qwen 3.6 35b NVFP4 — 256k ctx, 81 …
What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • Qwen 3.6 35b NVFP4 — 256k ctx, 81 tok/s • Qwen 3.6 27b NVFP4 — 256k ctx, 33 tok/s 2× DGX Spar
What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • Qwen 3.6 35b NVFP4 — 256k ctx, 81 tok/s • Qwen 3.6 27b NVFP4 — 256k ctx, 33 tok/s 2× DGX Sparks ← sweet spot! • DeepSeek v4 Flash — 1M ctx, 60 tok/s • MiMo-V2.5 — 1M ctx, Full Omni, 31 tok/s • Step-3.7-Flash — 256K ctx, 30 tok/s 3× DGX Sparks With three units, I recommend running a large model across two of them and a smaller one on the third. I'd personally run DeepSeek v4 Flash + Qwen 3.6 35B. This setup gives you DeepSeek v4 Flash speeds for coding, plus strong agentic workflows and image support from Qwen 3.6 35B. 4× DGX Sparks GLM 5.2 NVFP4 across all 4 units — if you have four units, this is the one to run! Links and repos below 👇
Related
- Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …
- pay attention anon. this is what local ai actually feels like in 2026. qwen 3.6 27b dense just knocked down the second test in my single fil…
- LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M…
- i don't think i need cloud models anymore
- NEW on Hugging Face: Hardware filters 🖥️ A new Hardware filter on the Models page results to models that fit a specific GPU, CPU, or Apple …
Source: Clem Delangue (X) | 2026-07-14