shipfeedAI news, curated daily

02:07:01 CET
26 JUL02:07:01shipfeed
pull to refreshlast sync
Just in — 30 new
§ tools · storyline

Adds Laguna on Apple GPUs via MLX engine

Ollama adds Laguna support on Apple GPUs via the MLX engine.

yesterday · · primary fetch1 sourceupdated yesterday ·

What's Changed Support Laguna on Apple GPUs via the MLX engine Quantize draft-model output heads at the requested type when creating speculative-decoding drafts. Fixed Qwen3 MoE decoding for differently-quantized experts, plus faster packed gate/up projection (~4–9% on M5 Max).

Full Changelog: https://github.com/ollama/ollama/compare/v0.32.3...v0.32.4

read full article on github.com
§ sources1 publication · timeline below
  1. github.comOllama v0.32.4primary