tag:github.com,2008:https://github.com/ollama/ollama/releases

Release notes from ollama

2026-09-16T21:06:08Z tag:github.com,2008:Repository/658928958/v0.34.2-rc1 2026-09-16T21:42:41Z

v0.34.2

<h2>What's Changed</h2> <ul> <li>llama.cpp updates</li> </ul> <p><strong>Full Changelog</strong>: <a class="commit-link" href="https://github.com/ollama/ollama/compare/v0.34.1...v0.34.2-rc0"><tt>v0.34.1...v0.34.2-rc0</tt></a></p> github-actions[bot] tag:github.com,2008:Repository/658928958/v0.34.2-rc0 2026-09-15T20:13:31Z

v0.34.2-rc0: llama.cpp: version bump b10969 (#18446)

<p>llama.cpp build changes resulted in duplicate symbols between libllama and libmtmd. This moves the compat patch into libllama with exported symbols.</p> dhiltgen tag:github.com,2008:Repository/658928958/v0.34.1 2026-09-15T20:10:42Z

v0.34.1

<h2>What's Changed</h2> <ul> <li>MLX safetensors <code>ollama create</code> no longer experimental. GGUF model creation now requires using llama.cpp tooling for safetensor conversion and quantization.</li> <li>Improved MLX memory handling on Apple Silicon</li> <li>Runaway repeat token detection now requires 100 repeat tokens for reduced false positives (e.g. OCR)</li> <li><code>/api/tags</code> is much faster on large model libraries (3.1 s → 294 ms cold in testing), and model capabilities are now reported consistently.</li> <li>Deprecated <code>typical_p</code>: it can no longer be set when creating new models, existing GGUF models retain support.</li> <li>MLX and llama.cpp updates</li> </ul> <p><strong>Full Changelog</strong>: <a class="commit-link" href="https://github.com/ollama/ollama/compare/v0.34.0...v0.34.1-rc1"><tt>v0.34.0...v0.34.1-rc1</tt></a></p> github-actions[bot] tag:github.com,2008:Repository/658928958/v0.34.1-rc2 2026-09-15T04:24:26Z

v0.34.1-rc2: API: Deprecate typical_p (#18448)

<p>Drop support for creating new models with typical_p parameters, while<br> retaining support for existing GGUF models with the setting.</p> dhiltgen tag:github.com,2008:Repository/658928958/v0.34.1-rc1 2026-09-14T20:34:03Z

v0.34.1-rc1

<p>mlx: add mlx patch to docker build context (<a class="issue-link js-issue-link" data-error-text="Failed to load title" data-id="5453819882" data-permission-text="Title is private" data-url="https://github.com/ollama/ollama/issues/18440" data-hovercard-type="pull_request" data-hovercard-url="/ollama/ollama/pull/18440/hovercard" href="https://github.com/ollama/ollama/pull/18440">#18440</a>)</p> dhiltgen tag:github.com,2008:Repository/658928958/v0.34.1-rc0 2026-09-14T16:49:46Z

v0.34.1-rc0: MLX: version bump (#18235)

<ul> <li> <p>MLX: version bump</p> </li> <li> <p>mlx: support ModelOpt global scales in MoE models</p> </li> <li> <p>address comments</p> </li> <li> <p>address comments</p> </li> </ul> dhiltgen tag:github.com,2008:Repository/658928958/v0.34.0 2026-09-10T06:18:35Z

v0.34.0

<h2>Use Ollama models in ChatGPT Desktop</h2> <p>Ollama models can now be used directly in ChatGPT Desktop, so you can keep your existing workflow while running open models. Setup is available from the Ollama app on MacOS.</p> <a target="_blank" rel="noopener noreferrer" href="https://private-user-images.githubusercontent.com/29360864/649245808-e10e299d-c11f-447d-9234-afa855824efe.png?jwt=eyJ0eXAiOiJKV1QiLCJhbGciOiJIUzI1NiJ9.eyJpc3MiOiJnaXRodWIuY29tIiwiYXVkIjoicmF3LmdpdGh1YnVzZXJjb250ZW50LmNvbSIsImtleSI6ImtleTUiLCJleHAiOjE3ODk2NDEyMDksIm5iZiI6MTc4OTY0MDkwOSwicGF0aCI6Ii8yOTM2MDg2NC82NDkyNDU4MDgtZTEwZTI5OWQtYzExZi00NDdkLTkyMzQtYWZhODU1ODI0ZWZlLnBuZz9YLUFtei1BbGdvcml0aG09QVdTNC1ITUFDLVNIQTI1NiZYLUFtei1DcmVkZW50aWFsPUFLSUFWQ09EWUxTQTUzUFFLNFpBJTJGMjAyNjA5MTclMkZ1cy1lYXN0LTElMkZzMyUyRmF3czRfcmVxdWVzdCZYLUFtei1EYXRlPTIwMjYwOTE3VDEwMjgyOVomWC1BbXotRXhwaXJlcz0zMDAmWC1BbXotU2lnbmF0dXJlPTE4ODU0NjdmZmU5MzcwNWNkNTJlOWMyZjQwMDg5MDc1MDhhMWVjZWQ3MmZlNzlhOTk0MzBmZTQwODE1YjMyM2YmWC1BbXotU2lnbmVkSGVhZGVycz1ob3N0JnJlc3BvbnNlLWNvbnRlbnQtdHlwZT1pbWFnZSUyRnBuZyJ9.7rdATRzoms8QgOoGzD1nm9ouQLb-JfUmsmfFRG089BU"><img width="1374" height="1300" alt="CleanShot 2026-09-08 at 11 04 07 AM@2x" src="https://private-user-images.githubusercontent.com/29360864/649245808-e10e299d-c11f-447d-9234-afa855824efe.png?jwt=eyJ0eXAiOiJKV1QiLCJhbGciOiJIUzI1NiJ9.eyJpc3MiOiJnaXRodWIuY29tIiwiYXVkIjoicmF3LmdpdGh1YnVzZXJjb250ZW50LmNvbSIsImtleSI6ImtleTUiLCJleHAiOjE3ODk2NDEyMDksIm5iZiI6MTc4OTY0MDkwOSwicGF0aCI6Ii8yOTM2MDg2NC82NDkyNDU4MDgtZTEwZTI5OWQtYzExZi00NDdkLTkyMzQtYWZhODU1ODI0ZWZlLnBuZz9YLUFtei1BbGdvcml0aG09QVdTNC1ITUFDLVNIQTI1NiZYLUFtei1DcmVkZW50aWFsPUFLSUFWQ09EWUxTQTUzUFFLNFpBJTJGMjAyNjA5MTclMkZ1cy1lYXN0LTElMkZzMyUyRmF3czRfcmVxdWVzdCZYLUFtei1EYXRlPTIwMjYwOTE3VDEwMjgyOVomWC1BbXotRXhwaXJlcz0zMDAmWC1BbXotU2lnbmF0dXJlPTE4ODU0NjdmZmU5MzcwNWNkNTJlOWMyZjQwMDg5MDc1MDhhMWVjZWQ3MmZlNzlhOTk0MzBmZTQwODE1YjMyM2YmWC1BbXotU2lnbmVkSGVhZGVycz1ob3N0JnJlc3BvbnNlLWNvbnRlbnQtdHlwZT1pbWFnZSUyRnBuZyJ9.7rdATRzoms8QgOoGzD1nm9ouQLb-JfUmsmfFRG089BU" content-type-secured-asset="image/png" style="max-width: 100%; height: auto; max-height: 1300px;"></a> <p>This release also improves structured output performance on Apple Silicon, adds support for OpenAI-compatible client tool search and response compaction.</p> <p><strong>Full Changelog</strong>: <a class="commit-link" href="https://github.com/ollama/ollama/compare/v0.33.3...v0.34.0"><tt>v0.33.3...v0.34.0</tt></a></p> github-actions[bot] tag:github.com,2008:Repository/658928958/v0.34.0-rc5 2026-09-09T21:51:50Z

v0.34.0-rc5

<p>openai: support standalone named function outputs (<a class="issue-link js-issue-link" data-error-text="Failed to load title" data-id="5404430682" data-permission-text="Title is private" data-url="https://github.com/ollama/ollama/issues/18348" data-hovercard-type="pull_request" data-hovercard-url="/ollama/ollama/pull/18348/hovercard" href="https://github.com/ollama/ollama/pull/18348">#18348</a>)</p> ParthSareen tag:github.com,2008:Repository/658928958/v0.34.0-rc4 2026-09-09T17:20:48Z

v0.34.0-rc4

<p>proxy: normalize namespaced commands in Full Access (<a class="issue-link js-issue-link" data-error-text="Failed to load title" data-id="5393403495" data-permission-text="Title is private" data-url="https://github.com/ollama/ollama/issues/18331" data-hovercard-type="pull_request" data-hovercard-url="/ollama/ollama/pull/18331/hovercard" href="https://github.com/ollama/ollama/pull/18331">#18331</a>)</p> ParthSareen tag:github.com,2008:Repository/658928958/v0.34.0-rc3 2026-09-09T00:09:48Z

v0.34.0-rc3

<p>openai: accept plaintext-labeled Codex agent messages (<a class="issue-link js-issue-link" data-error-text="Failed to load title" data-id="5393123059" data-permission-text="Title is private" data-url="https://github.com/ollama/ollama/issues/18329" data-hovercard-type="pull_request" data-hovercard-url="/ollama/ollama/pull/18329/hovercard" href="https://github.com/ollama/ollama/pull/18329">#18329</a>)</p> ParthSareen