v0.32.4
Mirrored from Ollama releases for archival readability. Support the source by reading on the original site.
What's Changed
- x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357
- test: harden flaky updater and transfer unit tests by @dhiltgen in #17378
- server: fix ps data race on scheduler loaded map by @dhiltgen in #17376
- qwen3_5: fix expert quantization handling and gather packed gate_up in one launch by @jessegross in #17336
- agent: permission skill loading by @ParthSareen in #17304
- cmd/tui: agent system prompt command by @ParthSareen in #17296
- mlx: keep loaded model memory resident by @dhiltgen in #17367
- x/create: quantize a draft model's output head at the requested type by @jessegross in #17383
- model: add Laguna MLX support by @dhiltgen in #17237
Full Changelog: v0.32.3...v0.32.4-rc0
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.