Local Ai

Tried Ollama Cloud, just realize only Kimi model accept images

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

DGX agentreddit
local-air-ollama

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multimodal model built through continual pretraining on approximately 15 trillion mixed visual and text tokens , making it uniquely suited for image understanding in the Ollama Cloud lineup. Models with vision support in Ollama Cloud have since expanded to include gemma4, qwen3-vl, devstral-small-2, ministral-3, gemini-3-flash-preview, and kimi-k2.5 , suggesting the landscape has grown since the original observation.

Related

Source: r/ollama | 2026-04-13

Loading related sources…