Ollama releases
92 articles archived · Visit source ↗ · RSS
-
Ollama releases dev-tools 1d ago
v0.32.10
agent: allow multiple edits per edit tool call ( #17711 )
15 -
Ollama releases dev-tools 1d ago
v0.32.10
What's Changed Models that don't set a repeat_penalty now default to 1.0 (off) instead of 1.1, matching other engines and speeding up speculative decoding; set a per-model parameter if an older model repeats itself. Faster prefill on NVFP4 MLX models with a global scale, about…
5 -
Ollama releases dev-tools 3d ago
v0.32.7
Muse Glimmer Note: Muse Glimmer is currently available via initial support via Ollama's MLX engine on Apple Silicon. Support for NVIDIA, AMD, and other platforms will be available in the coming days. Muse Glimmer , Meta's newest open model and the first released by Meta…
33 -
Ollama releases dev-tools 17d ago
v0.32.5
What's Changed mlx update by @dhiltgen in #17397 Full Changelog : v0.32.4...v0.32.5-rc0
20 -
Ollama releases dev-tools 19d ago
v0.32.4
What's Changed x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357 test: harden flaky updater and transfer unit tests by @dhiltgen in #17378 server: fix ps data race on scheduler loaded map by @dhiltgen in #17376 qwen3_5: fix expert quantization…
37 -
Ollama releases dev-tools 20d ago
v0.32.4-rc0: model: add Laguna MLX support (#17237)
model: add Laguna MLX support Add Laguna XS 2, XS 2.1, and S 2.1 support to the MLX model and create paths. Read the source config to apply one quantization policy across dense and routed MoE layers. Keep the tied output head and router at source precision, quantize supported…
19 -
Ollama releases dev-tools 21d ago
v0.32.3
What's Changed mlx update by @dhiltgen in #17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in #17250 docs: update retirements by @mxyng in #17289 model: align Laguna with upstream llama.cpp by @dhiltgen in #17335 Full Changelog : v0.32.2...v0.32.3-rc0
9 -
Ollama releases dev-tools 22d ago
v0.32.3
What's Changed mlx update by @dhiltgen in #17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in #17250 docs: update retirements by @mxyng in #17289 model: align Laguna with upstream llama.cpp by @dhiltgen in #17335 Full Changelog : v0.32.2...v0.32.3-rc0
11 -
Ollama releases dev-tools 23d ago
v0.32.2
What's Changed launch: keep Claude Code channels available by @hoyyeva in #17210 cmd: remove dead agent prompt wrappers by @ParthSareen in #17227 agent: reorder working directory instruction by @ParthSareen in #17228 agent: clean up semantics, UX, DX, and procedural code by…
24 -
Ollama releases dev-tools 23d ago
v0.32.2: test: revamp integration test entrpoints (#16560)
This refactors the existing integration tests into 3 priumary groups: fast, release, and library. It also refines some of the release tests to drop some of the older models and pick up newer models, while retaining the broad coverage in the library group.
25 -
Ollama releases dev-tools 23d ago
v0.32.2
What's Changed launch: keep Claude Code channels available by @hoyyeva in #17210 cmd: remove dead agent prompt wrappers by @ParthSareen in #17227 agent: reorder working directory instruction by @ParthSareen in #17228 agent: clean up semantics, UX, DX, and procedural code by…
32 -
Ollama releases dev-tools 24d ago
v0.32.2
What's Changed launch: keep Claude Code channels available by @hoyyeva in #17210 cmd: remove dead agent prompt wrappers by @ParthSareen in #17227 agent: reorder working directory instruction by @ParthSareen in #17228 agent: clean up semantics, UX, DX, and procedural code by…
31 -
Ollama releases dev-tools 29d ago
v0.32.1
What's Changed Improved Gemma 4 tool calling and multi-turn reasoning, including more reliable tool-response continuations Fixed a recurrent MLX model cache leak that could increase memory use across requests, and improved cache snapshot performance MLX text model loading now…
15 -
Ollama releases dev-tools 29d ago
v0.32.1-rc0
cmd: put current working dir in the system prompt ( #17188 )
5 -
Ollama releases dev-tools 1mo ago
v0.32.0
launch: rename Codex App integration to ChatGPT ( #17161 )
25 -
Ollama releases dev-tools 1mo ago
v0.31.2-rc1: create: harden GGUF create flows (#17062)
create: harden GGUF create flows lint
26 -
Ollama releases dev-tools 1mo ago
v0.31.1: mlx: tighten up gemma4 moe loading code (#16964)
This change allows .experts.gate_proj / .up_proj / .down_proj tensor names to each be used for both quantized (i.e. nvfp4 and mxfp8) and non-quantized (bf16) models. Previous to this only non-quantized models used that tensor naming scheme.
20 -
Ollama releases dev-tools 1mo ago
v0.31.0
launch: check for min version for hermes desktop ( #16912 )
4 -
Ollama releases dev-tools 1mo ago
v0.30.11
What's Changed launch: add thinking capability detection to opencode by @hoyyeva in #15434 launch: auto-install Claude Code by @hoyyeva in #16802 launch: auto-install opencode when missing by @hoyyeva in #16806 discover: fix inverted iGPU/dGPU Vulkan classification on Windows…
28 -
Ollama releases dev-tools 1mo ago
v0.30.11-rc1
parser/renderer: add Ornith 9B renderer/parser support ( #16920 )
23 -
Ollama releases dev-tools 1mo ago
v0.30.9-rc2
llm: context shift allow shiftable prompts ( #16764 )
31 -
Ollama releases dev-tools 1mo ago
v0.30.9-rc1
server: context shift for context windows larger than 8k, add error w…
28 -
Ollama releases dev-tools 2mo ago
v0.30.7
docs: update docs examples to use Gemma 4 instead of Gemma 3 ( #16607 )
7 -
Ollama releases dev-tools 2mo ago
v0.30.7-rc0
launch: use native Windows Hermes config path ( #16558 )
5