shipfeedAI news, curated daily

23:12:58 CET
7 AUG23:12:58shipfeed
pull to refreshlast sync
Just in — 8 new
§ topic

ollama

2 stories · 7d·6 sources covering·30 active storylines

Updated Tue, 04 Aug 2026 CEST·2 new storylines this week·live

What this is

Ollama is an open-source tool for running open-weight LLMs locally with one command. shipfeed tracks Ollama releases, new model support, and API and platform updates.

storylines this week30 active

Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc27

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc17

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Thursday, April 2, 2026’s edition
Ollama — Releases
OLLAMA · 1 source

Ollama v0.20.0

Gemma 4 Effective 2B (E2B) ``` ollama run gemma4:e2b ``` Effective 4B (E4B) ``` ollama run gemma4:e4b ``` 26B (Mixture of Experts model with 4B active parameters) ``` ollama run gemma4:26b ``` 31B (Dense) ``` ollama…

via github.com
Thursday, July 16, 2026’s edition
Ollama — Releases
OLLAMA · v0.32.1-rc0

Fixes MLX cache leak that could increase memory

What's Changed Improved Gemma 4 tool calling and multi-turn reasoning, including more reliable tool-response continuations Fixed a recurrent MLX model cache leak that could increase memory use across requests, and…

via github.com
Wednesday, June 3, 2026’s edition
Wednesday, May 13, 2026’s edition
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc31

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc21

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc15

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc20

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc29

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.30.0-rc22

This version of Ollama will change the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX is used to accelerate model inference on…

via github.com
Friday, March 27, 2026’s edition
Ollama — Releases
OLLAMA · 1 source

Ollama v0.19.0

Ollama is now powered by MLX on Apple Silicon in preview Ollama on Apple silicon is now built on top of Apple’s machine learning framework, MLX, to take advantage of its unified memory architecture…

via github.com
Tuesday, August 4, 2026’s edition
Ollama — Releases
OLLAMA · v0.32.6-rc0

Speeds up Qwen3.5 on Apple GPUs with speculative decoding

What's Changed Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for speculative decoding automatically `/v1/chat/completions` streaming now matches OpenAI's wire format: `role` only on the…

via github.com
Saturday, July 25, 2026’s edition
Ollama — Releases
OLLAMA · v0.32.4

Adds Laguna on Apple GPUs via MLX engine

What's Changed Support Laguna on Apple GPUs via the MLX engine Quantize draft-model output heads at the requested type when creating speculative-decoding drafts. Fixed Qwen3 MoE decoding for differently-quantized…

via github.com
Thursday, July 23, 2026’s edition
Ollama — Releases
OLLAMA · v0.32.3-rc0

Fixes incomplete GLM tool calls

What's Changed mlx update by @dhiltgen in https://github.com/ollama/ollama/pull/17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in https://github.com/ollama/ollama/pull/17250 docs: update…

via github.com
Monday, July 20, 2026’s edition
Ollama — Releases
OLLAMA · v0.32.2

Adds agent skills system

What's Changed launch: keep Claude Code channels available by @hoyyeva in https://github.com/ollama/ollama/pull/17210 cmd: remove dead agent prompt wrappers by @ParthSareen in…

via github.com
Saturday, July 11, 2026’s edition
Ollama — Releases
OLLAMA · v0.32.0

Adds agent UI and parser for Qwen3.5/Next models

What's Changed create: select the qwen3.5 parser and renderer for Qwen3.5/Next by @jessegross in https://github.com/ollama/ollama/pull/17078 launch: warn before old agent models by @ParthSareen in…

via github.com
Sunday, June 28, 2026’s edition
Thursday, June 25, 2026’s edition
Wednesday, June 3, 2026’s edition
Ollama — Releases
OLLAMA · v0.30.3

Adds support for Gemma4-12B models

What's Changed models: add support for gemma4-12b by @pdevine in https://github.com/ollama/ollama/pull/16457 Full Changelog: https://github.com/ollama/ollama/compare/v0.30.2...v0.30.3

via github.com
Tuesday, June 2, 2026’s edition
Thursday, May 14, 2026’s edition
Ollama — Releases
OLLAMA · 1 source

Ollama v0.24.0-rc1

What's Changed mlx: add memory trace logging by @dhiltgen in https://github.com/ollama/ollama/pull/16131 launch: codex app integration by @ParthSareen in https://github.com/ollama/ollama/pull/16120 Full Changelog…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.24.0

What's Changed mlx: add memory trace logging by @dhiltgen in https://github.com/ollama/ollama/pull/16131 launch: codex app integration by @ParthSareen in https://github.com/ollama/ollama/pull/16120 Full Changelog…

via github.com
Ollama — Releases
OLLAMA · 1 source

Ollama v0.24.0-rc0

What's Changed mlx: add memory trace logging by @dhiltgen in https://github.com/ollama/ollama/pull/16131 launch: codex app integration by @ParthSareen in https://github.com/ollama/ollama/pull/16120 Full Changelog…

via github.com
Tuesday, May 5, 2026’s edition
Ollama — Releases
OLLAMA · 1 source

Ollama v0.23.1

Gemma 4 MTP (Multi-token Processing) for the MLX runner Gemma 4 MTP speculative decoding is now supported on Macs. This can give over a 2x speed increase for the Gemma 4 31B model on coding tasks. ``` ollama run…

via github.com