Local Ai
A caveman qwen3.6 27B
Just saw this on huggingface: https://huggingface.co/ProCreations/grug-27b The benchmarks claim that it's quite a bit better than qwen3.6 27B original and that they reduced the amount of necessary tok
Just saw this on huggingface: https://huggingface.co/ProCreations/grug-27b The benchmarks claim that it's quite a bit better than qwen3.6 27B original and that they reduced the amount of necessary tokens by more than 90%. It would make 27B running on my old laptop at 3tps feel more like 30tps for the thinking part, if true. Couldn't test it yet. submitted by /u/AppealSame4367 [link] [comments]
Related
- Bonsai 27B: 1-bit dense LLM running locally in your browser using custom WebGPU kernels
- Absurd claim: the distilled model outperforms the originals
- Is hosting an SLM cheaper than using APIs with 20$ cost per user?
- I bundled a fully local LLM inside my Unity game. No internet, no cloud, no API key. The conversation is the gameplay.
- Cost Analysis of my $6.4k Local LLM Server
Source: r/LocalLLaMA | 2026-07-23