Model Releases
model: support Longcat-Flash (need testing) by ngxson · Pull Request #19182 · ggml-org/llama.cpp
This PR should be ready for testing now. I tested with a very small (8B params) sub-model extracted from the original one. Appreciate if someone can test with the bigger model. GGUF(for testing) from
This PR should be ready for testing now. I tested with a very small (8B params) sub-model extracted from the original one. Appreciate if someone can test with the bigger model. GGUF(for testing) from PR: (Please check latest comments at bottom for updated GGUFs) https://huggingface.co/ggml-org/LongCat-Flash-Chat-GGUF/tree/main submitted by /u/pmttyji [link] [comments]
Related
- Support Step3.5/3.7 flash mtp3 by forforever73 · Pull Request #24340 · ggml-org/llama.cpp
- Uncensored Multi-Model Releases, LongCat-Flash-Lite with MTPs, Jamba2-Mini, Qwen3.5-9B-Nikusui-v1 with MTPs and Qwen3.5-27B-Nikusui-v1 with MTPs, Available in Safetensors and GGUF Formats!
- DeepSeek-V4-Flash-0731 unsloth gguf on A100
Source: r/LocalLLaMA | 2026-08-08