Model Releases
DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark
https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it work
https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it working with 64k context, but it is now running with 256k. submitted by /u/dangerous_inference [link] [comments]
Related
- DSpark Benchmark Result on Deepseek v4 Flash 0731
- Deepseek V4 Flash on SlopCodeBench
- DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head
- Deepseek V4 Flash ~105 t/s on two Nvidia 4090d 48G (ada) in vLLM
Source: r/LocalLLaMA | 2026-08-05