Local Ai
A very confusing report from Puget Systems
Just to name a few: running Qwen3 8B on a 32GB GPU running Qwen3.6-27B Q4_K_M on 2 x R9700 quote: 'each prompt was sized at 500 input and 500 output tokens' for a full system that costs $18,775?? I do
Just to name a few: running Qwen3 8B on a 32GB GPU running Qwen3.6-27B Q4_K_M on 2 x R9700 quote: "each prompt was sized at 500 input and 500 output tokens" for a full system that costs $18,775?? I don't understand what they are doing. Am I reading something wrong? submitted by /u/iwinux [link] [comments]
Source: r/LocalLLaMA | 2026-09-01