Model Releases

DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark

https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it work

DGX agentreddit
model-releasesr-localllama

https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it working with 64k context, but it is now running with 256k. submitted by /u/dangerous_inference [link] [comments]

Related

Source: r/LocalLLaMA | 2026-08-05

Loading related sources…