shipfeedAI news, curated daily

23:04:39 CET
25 AUG23:04:39shipfeed
pull to refreshlast sync
Just in — 8 new
§ topic

llama

0 stories · 7d·6 sources covering·5 active storylines

Updated Fri, 07 Aug 2026 CEST·0 new storylines this week·live

What this is

Llama is Meta's family of open-weight large language models. shipfeed tracks every Llama release, new generation, weight drop, and the ecosystem of fine-tunes built on it.

storylines this week5 active

Friday, August 7, 2026’s edition
Friday, May 29, 2026’s edition
Saturday, April 20, 2024’s edition
Monday, June 1, 2026’s edition
llama.cpp — Releases
LLAMA.CPP · b9460

Limits max outputs of llama_context to save VRAM

llama: limit max outputs of `llama_context` (#23861) llama: save more VRAM by reserving n_outputs == n_seqs when possible add n_outputs_per_seq move n_outputs_max to server-context change ubatch to batch everywhere…

via github.com