Model Releases
Another day and another full frontier model running on your computer. Been teething DeepSeek V4 Flash on over 60 employees at The Zero Human…
Another day and another full frontier model running on your computer. Been teething DeepSeek V4 Flash on over 60 employees at The Zero Human Company for a few hours and it is stunning. Between Kimi K3
Another day and another full frontier model running on your computer. Been teething DeepSeek V4 Flash on over 60 employees at The Zero Human Company for a few hours and it is stunning. Between Kimi K3 and this nearly 80% of most folks needs is covered locally. https://huggingface.co/ox-ox/DeepSeek-V4-Flash-0731-GGUF
Related
- DeepSeek v4 Flash with local inference after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST t…
- Model is available here @simonw @ivanfioravanti https://huggingface.co/mlx-community/DeepSeek-V4-Flash-2bit-DQ
- I guarantee you are sleeping on small models. Deepseek V4 Flash can do ~80% of the tasks you ask Claude or Codex for. It is 137x cheaper per…
Source: Clem Delangue (X) | 2026-08-02