# Ollama v0.33.1 - Product: Ollama (https://whatsnew.fyi/product/ollama) - Vendor: Ollama - Date: 2026-08-26 - Version: v0.33.1 - Original notes: https://github.com/ollama/ollama/releases/tag/v0.33.1 - Permalink: https://whatsnew.fyi/product/ollama/releases/v0.33.1-rc1 What's New is an index, not a publisher: every entry below links to the vendor's own release notes, which are the authoritative source. Entries are labelled where they are hand-curated sample data, pre-releases, or drawn from a secondary source such as a developer blog. Reuse: the summaries, labels and curation here are © What's New. Quote freely with attribution and a link back; wholesale republication of the corpus is not permitted — terms: https://whatsnew.fyi/terms. The vendors' own release notes remain their publishers'. --- - **added** — Add Qwen3.8 Flash Next support for MLX - **changed** — Make external compat patches idempotent in cmake - **changed** — Update MLX and llama.cpp - **added** — Add structured output support to mlxrunner - **fixed** — Avoid Metal GPU timeouts when loading models from slow storage in mlxrunner ##### What's Changed * MLX: Qwen3.8 Flash Next support * cmake: make external compat patches idempotent * MLX and llama.cpp update * mlxrunner: add structured output support * mlxrunner: avoid Metal GPU timeouts when loading models from slow storage ##### New Contributors * @pd95 made their first contribution in https://github.com/ollama/ollama/pull/17948 **Full Changelog**: https://github.com/ollama/ollama/compare/v0.33.0...v0.33.1