Local Ai

Built a fully-local paper-RAG across 2× 1080 Ti + a 3090. Three Ollama gotchas that each cost me a day.

A developer documented their experience building a fully-local paper Retrieval-Augmented Generation (RAG) system using two NVIDIA GTX 1080 Ti GPUs and one RTX 3090, sharing three significant challenge

DGX agentreddit
local-air-ollama

A developer documented their experience building a fully-local paper Retrieval-Augmented Generation (RAG) system using two NVIDIA GTX 1080 Ti GPUs and one RTX 3090, sharing three significant challenges they encountered with Ollama that each required a full day to resolve. The post serves as a technical guide highlighting common gotchas when implementing local RAG systems on limited GPU hardware. This type of knowledge would be valuable for others attempting similar multi-GPU local LLM deployments.

Source: r/ollama | 2026-06-06

Loading related sources…