Model Releases
Appreciation for Gemma 4 26b A4b
I really love this model, I have been using the q4_k_l by Bartowski (I have heard QAT is quite the downgrade in some aspects) and it handles every task I throw at it easily. Agentic and coding perform
I really love this model, I have been using the q4_k_l by Bartowski (I have heard QAT is quite the downgrade in some aspects) and it handles every task I throw at it easily. Agentic and coding performance is not as good as Qwen of course but good enough and I have found myself to be constantly surprised by it considering the size and speed. I really like its personality. It's a great writer and soulful especially with lower soft logit capping values. Of course it also has native multimodality which is a big plus. But the real star of the show is its language capabilities and knowledge. It is excellent in German so it feels like a big cloud model in that regard and it knows quite a lot for such a speedy local model, too. Definately more world knowledge than Qwen. And all of that runs at around 10-23 token/s on my aging laptop with 600 token/s prefill performance. What is your experience with this model? If you haven't been using it for a while, consider giving it a go again with the new chat template. submitted by /u/dampflokfreund [link] [comments]
Related
- Current smallest usable coding model
- Cactus Hybrid: We taught Gemma 4 to know when it's wrong
- Gemma 4 Chat Template now has preserve thinking
- Anybody else noticing how good gemma-4-26b-a4b is with one-shotting three.js?
Source: r/LocalLLaMA | 2026-07-28