Model Releases
DFlash 2 available for Qwen 3.8 27B and Muse Glimmer
Apparently a second version of DFlash from the original authors of DFlash GGUF quants are already made available with an accompanying llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/27342 sub
Apparently a second version of DFlash from the original authors of DFlash GGUF quants are already made available with an accompanying llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/27342 submitted by /u/rerri [link] [comments]
Related
- model: support Longcat-Flash (need testing) by ngxson · Pull Request #19182 · ggml-org/llama.cpp
- Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)
- New Muse-Glimmer-30B SoTA Quants - hopefully a new lineup :)
- Support Step3.5/3.7 flash mtp3 by forforever73 · Pull Request #24340 · ggml-org/llama.cpp
Source: r/LocalLLaMA | 2026-08-18