Model Releases

Appreciation for Gemma 4 26b A4b

I really love this model, I have been using the q4_k_l by Bartowski (I have heard QAT is quite the downgrade in some aspects) and it handles every task I throw at it easily. Agentic and coding perform

DGX agentreddit
model-releasesr-localllama

I really love this model, I have been using the q4_k_l by Bartowski (I have heard QAT is quite the downgrade in some aspects) and it handles every task I throw at it easily. Agentic and coding performance is not as good as Qwen of course but good enough and I have found myself to be constantly surprised by it considering the size and speed. I really like its personality. It's a great writer and soulful especially with lower soft logit capping values. Of course it also has native multimodality which is a big plus. But the real star of the show is its language capabilities and knowledge. It is excellent in German so it feels like a big cloud model in that regard and it knows quite a lot for such a speedy local model, too. Definately more world knowledge than Qwen. And all of that runs at around 10-23 token/s on my aging laptop with 600 token/s prefill performance. What is your experience with this model? If you haven't been using it for a while, consider giving it a go again with the new chat template. submitted by /u/dampflokfreund [link] [comments]

Related

Source: r/LocalLLaMA | 2026-07-28

Loading related sources…