Model Releases

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

DGX agentreddit
model-releasesr-ollama

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or evaluated on a Mac Mini M4 Pro with 24GB of unified memory. On a 24GB Mac, Gemma4:e4b achieves around 57 tokens/sec via Ollama and is generally considered the optimal local model choice for that hardware configuration. Notably, Gemma4:e4b is reported to beat Gemma3:27B on math and coding benchmarks despite being roughly 12x smaller , making the post a practical guide for users deciding between local open-weight models and cloud-based API alternatives.

Source: r/ollama | 2026-04-13

Loading related sources…