Local Ai

Running gpt-oss:20b locally and grading it head to head against a frontier model on real tasks. It held up better than I expected

I serve a free local model on my Mac Mini and route real agent work to it. To check I was not fooling myself, I set up a blind grader that replays frontier tasks locally and scores both. https://previ

DGX agentreddit
local-air-ollama

I serve a free local model on my Mac Mini and route real agent work to it. To check I was not fooling myself, I set up a blind grader that replays frontier tasks locally and scores both. https://preview.redd.it/u17xgqrdo7hh1.png?width=3594&format=png&auto=webp&s=06fdc4fe9285a1d911b981d39c034b50e698ba64 10 blind rematches with a mean gap -0.05, the local gpt-oss-20b model won 4, and one writing task went local 3.70 vs frontier 2.80. It lost on the dense synthesis, so my for now my work stays remote. For everything else, local is a wash or better at no per call cost. Happy to share the model, the serving config, and how the grader works. submitted by /u/AIForOver50Plus [link] [comments]

Source: r/ollama | 2026-08-03

Loading related sources…