Model Releases
Current smallest usable coding model
I've been seeing a lot of news about the latest gemma 4 and qwen 3.6 being really good and the current go-to models but those are out of reach for my GPU at the moment. With 4GB VRAM and 40 GB RAM, I
I've been seeing a lot of news about the latest gemma 4 and qwen 3.6 being really good and the current go-to models but those are out of reach for my GPU at the moment. With 4GB VRAM and 40 GB RAM, I was wondering which other smaller model is the best for agentic coding and also if I should stick with llama.cpp for running it or not? submitted by /u/GamerWael [link] [comments]
Source: r/LocalLLaMA | 2026-07-27