Local Ai
v0.32.4
What's Changed x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357 test: harden flaky updater and transfer unit tests by @dhiltgen in #17378 server: fix ps data race o
What's Changed x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357 test: harden flaky updater and transfer unit tests by @dhiltgen in #17378 server: fix ps data race on scheduler loaded map by @dhiltgen in #17376 qwen3_5: fix expert quantization handling and gather packed gate_up in one launch by @jessegross in #17336 agent: permission skill loading by @ParthSareen in #17304 cmd/tui: agent system prompt command by @ParthSareen in #17296 mlx: keep loaded model memory resident by @dhiltgen in #17367 x/create: quantize a draft model's output head at the requested type by @jessegross in #17383 model: add Laguna MLX support by @dhiltgen in #17237 Full Changelog: v0.32.3...v0.32.4-rc0
Related
- v0.32.2-rc1: server: detect download stalls before the first byte (#17259)
- v0.32.2-rc3: test: revamp integration test entrpoints (#16560)
- v0.32.0-rc0
Source: Ollama Releases | 2026-07-25