Model Releases
DeepSeek V4 Flash with Antirez Dwarfstar 4 is amazing.
Note: I use a Mac Studio M3U with 512 GB RAM, so this is not for everyone. I have been using antirez/ds4 with DS4 Flash for a few weeks now. Top quality, I am really impressed. This combination just d
Note: I use a Mac Studio M3U with 512 GB RAM, so this is not for everyone. I have been using antirez/ds4 with DS4 Flash for a few weeks now. Top quality, I am really impressed. This combination just delivers. It is incomparable to Qwen 3.6 27B or GLM 5.2 (which I can only run heavily quantized at Q3). But no need to rely on Sol or Opus anymore to solve difficult problems. On average, I get 35 tok/sec. But requires 390 GB of RAM for full quality... If you have a Mac with at least 256 GB, you should try this. Big thanks to antirez for this amazing framework! submitted by /u/pj-frey [link] [comments]
Related
- DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head
- Extremely slow DSpark draft model performance (1-2 t/s) with DeepSeek-V4-Flash on llama-server compared to MTP?
Source: r/LocalLLaMA | 2026-08-17