Ollama releases
139 articles archived · Visit source ↗ · RSS
-
Ollama releases dev-tools 4d ago
v0.34.4-rc1: mlxrunner: Update XGrammar to 0.2.7 for structured outputs
We pick up schema fixes for typed dictionary values and short arrays.
12 -
Ollama releases dev-tools 4d ago
v0.34.4: mlxrunner: Update XGrammar to 0.2.7 for structured outputs
We pick up schema fixes for typed dictionary values and short arrays.
18 -
Ollama releases dev-tools 9d ago
v0.34.3-rc1
server: allow registry cross-host redirects among allowlisted hosts (…
25 -
Ollama releases dev-tools 9d ago
v0.34.3
server: allow registry cross-host redirects among allowlisted hosts (…
16 -
Ollama releases dev-tools 9d ago
v0.34.3-rc0
api: expose model thinking levels and defaults ( #18473 )
35 -
Ollama releases dev-tools 10d ago
v0.34.2
x/transfer, server: tighten redirect handling for registry requests (…
19 -
Ollama releases dev-tools 10d ago
v0.34.2
What's Changed llama.cpp updates Full Changelog : v0.34.1...v0.34.2-rc0
19 -
Ollama releases dev-tools 11d ago
v0.34.2
What's Changed llama.cpp updates Full Changelog : v0.34.1...v0.34.2-rc0
4 -
Ollama releases dev-tools 13d ago
v0.34.1-rc1
mlx: add mlx patch to docker build context ( #18440 )
14 -
Ollama releases dev-tools 13d ago
v0.34.1-rc0: MLX: version bump (#18235)
MLX: version bump mlx: support ModelOpt global scales in MoE models address comments address comments
19 -
Ollama releases dev-tools 18d ago
v0.34.0
Use Ollama models in ChatGPT Desktop Ollama models can now be used directly in ChatGPT Desktop, so you can keep your existing workflow while running open models. Setup is available from the Ollama app on MacOS. This release also improves structured output performance on Apple…
34 -
Ollama releases dev-tools 18d ago
v0.34.0-rc5
openai: support standalone named function outputs ( #18348 )
12 -
Ollama releases dev-tools 18d ago
v0.34.0
Use Ollama models in ChatGPT Desktop Ollama models can now be used directly in ChatGPT Desktop, so you can keep your existing workflow while running open models. Setup is available from the Ollama app on MacOS. This release also improves structured output performance on Apple…
18 -
Ollama releases dev-tools 19d ago
v0.34.0
Use Ollama models in ChatGPT Desktop Ollama models can now be used directly in ChatGPT Desktop, so you can keep your existing workflow while running open models. Setup is available from the Ollama app on MacOS. This release also improves structured output performance on Apple…
33 -
Ollama releases dev-tools 19d ago
v0.34.0-rc2
openai: finalize responses at the web search limit ( #18328 )
15 -
Ollama releases dev-tools 23d ago
v0.34.0-rc1
app: harden Codex desktop proxy handling ( #18244 )
30 -
Ollama releases dev-tools 25d ago
v0.33.3: gemma4: image and audio input support
Safetensors gemma4 imports served by the MLX engine now answer image and audio chats. Images run through both vision architectures: the transformer tower (26B, 31B, e-series) and the 12B's encoder-free unified embedder. Audio arrives through the same intake the ollama API…
22 -
Ollama releases dev-tools 26d ago
v0.33.3-rc0: llama.cpp: version bump b10729 (#18160)
llama.cpp: version bump b10729 Regenerate the compat hooks patch for b10729: upstream removed the whole-tensor load_data_for read (last consumer was llama-quantize, which now reads slabs via load_data_range). Keep the existing hook surface (constructor, skip loops,…
32 -
Ollama releases dev-tools 1mo ago
v0.33.2-rc1
app: list account cloud models for Claude ( #18077 )
16 -
Ollama releases dev-tools 1mo ago
v0.33.1
What's Changed MLX: Qwen3.8 Flash Next support cmake: make external compat patches idempotent MLX and llama.cpp update mlxrunner: add structured output support mlxrunner: avoid Metal GPU timeouts when loading models from slow storage New Contributors @pd95 made their first…
6 -
Ollama releases dev-tools 1mo ago
v0.33.1-rc1: linux: fix llama.cpp docker build (#18040)
Build context was missing the new cmake common utility.
29 -
Ollama releases dev-tools 1mo ago
v0.33.1-rc0: MLX: Qwen3.8 Flash Next support (#18032)
MLX: Qwen3.8 Flash Next support review comments
25 -
Ollama releases dev-tools 1mo ago
v0.33.0
What's Changed Claude Desktop Turn individual Ollama models on or off for use in Claude, directly from the menu bar Choose from your available Ollama models from within Claude; cloud models appear only when you're signed in A new Apps view manages app integrations with copyable…
6 -
Ollama releases dev-tools 1mo ago
v0.33.0
proxy: preserve string content during image fallback ( #18002 )
20 -
Ollama releases dev-tools 1mo ago
v0.33.0
What's Changed Claude Desktop Turn individual Ollama models on or off for use in Claude, directly from the menu bar Choose from your available Ollama models from within Claude; cloud models appear only when you're signed in A new Apps view manages app integrations with copyable…
4 -
Ollama releases dev-tools 1mo ago
v0.33.0
What's Changed mlx: fix mac assumptions on linux/windows by @dhiltgen in #17898 mlx update by @dhiltgen in #17886 lint fixes by @dhiltgen in #17897 app: add claude desktop app by @ParthSareen in #17899 app: polish onboarding layout and disable zoom by @hoyyeva in #17885 launch:…
20 -
Ollama releases dev-tools 1mo ago
v0.33.0
What's Changed mlx: fix mac assumptions on linux/windows by @dhiltgen in #17898 mlx update by @dhiltgen in #17886 lint fixes by @dhiltgen in #17897 app: add claude desktop app by @ParthSareen in #17899 app: polish onboarding layout and disable zoom by @hoyyeva in #17885 launch:…
31 -
Ollama releases dev-tools 1mo ago
v0.32.15-rc1: ci: plumb temporary MLX patch through to docker stages (#17874)
Follow up to #17850
35 -
Ollama releases dev-tools 1mo ago
v0.32.15-rc0: mlx update (#17850)
Temporarily carry ml-explore/mlx-c#127
28 -
Ollama releases dev-tools 1mo ago
v0.32.13: qwen3.8: support developer instructions (#17749)
qwen3.8: support developer instructions Qwen3.8 does not define a developer role, while OpenAI-compatible coding agents commonly send developer instructions before user messages. Fold the leading system/developer instruction prefix into a single system turn before Qwen3.8…
35 -
Ollama releases dev-tools 1mo ago
v0.32.12: qwen3.8: add renderer and MLX import support
Qwen3.8 keeps the Qwen3.5 model architecture and parser, but its chat template adds reasoning-effort and preserved-thinking semantics. Detect those template markers during safetensors import, select the qwen3.8 renderer, and cover thinking, tools, continuation, and malformed…
14 -
Ollama releases dev-tools 1mo ago
v0.32.10
agent: allow multiple edits per edit tool call ( #17711 )
15 -
Ollama releases dev-tools 1mo ago
v0.32.10
What's Changed Models that don't set a repeat_penalty now default to 1.0 (off) instead of 1.1, matching other engines and speeding up speculative decoding; set a per-model parameter if an older model repeats itself. Faster prefill on NVFP4 MLX models with a global scale, about…
5