Local Ai

v0.32.4

What's Changed x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357 test: harden flaky updater and transfer unit tests by @dhiltgen in #17378 server: fix ps data race o

DGX agentgithub
local-aiollama-releases

What's Changed x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357 test: harden flaky updater and transfer unit tests by @dhiltgen in #17378 server: fix ps data race on scheduler loaded map by @dhiltgen in #17376 qwen3_5: fix expert quantization handling and gather packed gate_up in one launch by @jessegross in #17336 agent: permission skill loading by @ParthSareen in #17304 cmd/tui: agent system prompt command by @ParthSareen in #17296 mlx: keep loaded model memory resident by @dhiltgen in #17367 x/create: quantize a draft model's output head at the requested type by @jessegross in #17383 model: add Laguna MLX support by @dhiltgen in #17237 Full Changelog: v0.32.3...v0.32.4-rc0

Related

Source: Ollama Releases | 2026-07-25

Loading related sources…