Fixes MLX cache leak that could increase memory
Ollama releases v0.32.1-rc0 fixing a recurrent MLX cache leak and improving Gemma 4 tool calling.
What's Changed Improved Gemma 4 tool calling and multi-turn reasoning, including more reliable tool-response continuations Fixed a recurrent MLX model cache leak that could increase memory use across requests, and improved cache snapshot performance MLX text model loading now respects `OLLAMA_LOAD_TIMEOUT` Agent web search and fetch now tell users to run `ollama signin` when authentication is required The interactive agent now receives the current working directory for better project context Fixed `ollama launch` so choosing Pick another model for a deprecated model passed with `--model` opens the model picker Updated VS Code setup documentation for the official Ollama extension Full Changelog: https://github.com/ollama/ollama/compare/v0.32.0...v0.32.1-rc0
- github.comOllama v0.32.1-rc0primary
- github.comOllama v0.32.1
§ how this story moved
- primary — Ollama — Releases publishes the launch post.
- Ollama — Releases picks up coverage.