shipfeedAI news, curated daily

00:21:28 CET
28 SEPT00:21:28shipfeed⋯
pull to refreshlast sync
Just in — 8 new
§ topic

local-llm

87 stories · 7d·6 sources covering·30 active storylines

Updated Fri, 25 Sept 2026 CEST·87 new storylines this week·live

What this is

Local LLMs are open-weight models you can run on your own hardware. shipfeed tracks open-weight releases, quantization, and local inference tools like llama.cpp and Ollama.

storylines this week30 active

llama.cpp — Releases
LLAMA.CPP · b9496

Fixes Gemma 4 unified FPE on mtmd

mtmd: fix Gemma 4 unified FPE (#24088) macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu arm64 (CPU) Ubuntu…

via github.com
Monday, August 10, 2026’s edition
Tuesday, September 22, 2026’s edition
Wednesday, August 19, 2026’s edition
Monday, August 10, 2026’s edition
Monday, July 27, 2026’s edition
Tuesday, July 21, 2026’s edition
Friday, July 17, 2026’s edition
Thursday, July 16, 2026’s edition
Wednesday, June 10, 2026’s edition
Monday, June 1, 2026’s edition
Friday, April 3, 2026’s edition
Smol AI — Daily
AI · 1 source

not much happened today

Gemma 4 was launched by Google under an Apache 2.0 license, marking a significant open-model release focused on reasoning, agentic workflows, multimodality, and on-device use. It outperforms models 10x larger and has…

via news.smol.ai
Wednesday, March 11, 2026’s edition
Smol AI — Daily
AI · 1 source

not much happened today

NVIDIA’s Nemotron 3 Super is a 120B parameter / ~12B active open model featuring a hybrid Mamba-Transformer / SSM Latent MoE architecture and 1M context window, delivering up to 2.2x faster inference than GPT-OSS-120B…

via news.smol.ai
Friday, September 25, 2026’s edition
Ollama — Releases
OLLAMA · v0.40.0-rc0

Runs models on MLX by default on Apple Silicon

What's Changed Models run on MLX on Apple Silicon by default In this release, model architectures supported by the MLX runner will run by default on Apple Silicon devices. ``` ollama pull qwen3.8 ollama run qwen3.8 ```…

via github.com
Thursday, September 24, 2026’s edition
Wednesday, September 23, 2026’s edition
llama.cpp — Releases
LLAMA.CPP · v0.5.0

Speeds up CUDA conv2d with implicit GEMM

Overview This release focuses on backend performance and correctness, broader model coverage, and more robust server/router operation. It adds HRM-Text (DFM Mimir 1B) support, MiMo-V2.6 and HunyuanOCR conversion…

via github.com
Tuesday, September 22, 2026’s edition
Monday, August 31, 2026’s edition
Friday, August 28, 2026’s edition
HN Algolia — Front-Page AI
AI · 1 source

GLM-5.3 is now open-weight

https://twitter.com/Zai_org/status/2093354097122455713https://z.ai/blog/glm-5.3

via huggingface.co
Thursday, August 27, 2026’s edition
Wednesday, August 26, 2026’s edition
Tuesday, August 25, 2026’s edition
llama.cpp — Releases
LLAMA.CPP · v0.3.0

Adds dots3-note multimodal model with DSA-ISWA KV cache

## Overview llama.cpp 0.3.0 introduces the dots3-note multimodal model (with a new DSA-ISWA KV cache), MTP support for GLM-4.5-Air, and tensor-split (`-sm tensor`) plus multi-sequence rollback fixes for DeepSeek 4…

via github.com
Sunday, August 23, 2026’s edition
Wednesday, August 19, 2026’s edition
Monday, August 17, 2026’s edition
Friday, August 14, 2026’s edition
Wednesday, August 12, 2026’s edition
Tuesday, August 11, 2026’s edition
Monday, August 10, 2026’s edition
local-llm — shipfeed