Local Ai
llama4 108b
This Reddit thread on r/ollama discusses running Meta's Llama 4 Maverick — a ~108B parameter model — locally using Ollama. Llama 4 models are natively multimodal AI models supporting text and image un
This Reddit thread on r/ollama discusses running Meta's Llama 4 Maverick — a ~108B parameter model — locally using Ollama. Llama 4 models are natively multimodal AI models supporting text and image understanding, leveraging a mixture-of-experts (MoE) architecture for high performance. The Llama 4 Maverick variant specifically uses 17 billion active parameters with 128 experts, and is available on Ollama at approximately 245GB with a 1M token context window, supporting both text and image input.
Related
- Recommended Model for a 4060ti 8gb and 16gb ram
- Using Ollama Gemma4 models via OpenWebUI on my phone and it’s been a good experience
- What's model should I run?
- Has anyone actually gotten a reliable local AI system running?
- Tried running LLMs locally to save API costs… ended up waiting 13 minutes for ONE response 🤡
- Any models?
Source: r/ollama | 2026-04-13