shipfeedAI news, curated daily

01:53:27 CET
12 OCT01:53:27shipfeed⋯
pull to refreshlast sync
Just in — 8 new
§ topic

llama

0 stories · 7d·6 sources covering·6 active storylines

Updated Mon, 21 Sept 2026 CEST·0 new storylines this week·live

What this is

Llama is Meta's family of open-weight large language models. shipfeed tracks every Llama release, new generation, weight drop, and the ecosystem of fine-tunes built on it.

storylines this week6 active

Friday, August 7, 2026’s edition
Friday, May 29, 2026’s edition
Saturday, April 20, 2024’s edition
Monday, September 21, 2026’s edition
Monday, June 1, 2026’s edition
llama.cpp — Releases
LLAMA.CPP · b9460

Limits max outputs of llama_context to save VRAM

llama: limit max outputs of `llama_context` (#23861) llama: save more VRAM by reserving n_outputs == n_seqs when possible add n_outputs_per_seq move n_outputs_max to server-context change ubatch to batch everywhere…

via github.com