Hardware
IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7
This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The
This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The RTX PRO 4000 Blackwell is a professional-grade workstation GPU built on NVIDIA's Blackwell architecture, featuring GDDR7 memory and 5th-generation Tensor Cores that accelerate local LLM inference. With its VRAM capacity, the card is well-suited for running small-to-medium parameter models (up to ~20B parameters) fully in GPU memory, enabling fast, private, on-device AI inference without relying on cloud services.
Related
- Is an nvidia DGK Spark or similar worth it?
- any decent model to run on 9070xt locally
- Recommended Model for a 4060ti 8gb and 16gb ram
- Tried running LLMs locally to save API costs… ended up waiting 13 minutes for ONE response 🤡
- Any models?
- What's model should I run?
Source: r/ollama | 2026-04-14