llama.cpp v0.4.1 · a signature changes, and GDN normalisation is fixed — from issue No. 031nofeed.dev
Issue No. 031 · 15 September 2026
Breaking

llama.cpp v0.4.1 · a signature changes, and GDN normalisation is fixed

Published 18:27 UTC 14 Sep. llama_sampler_chain_n() returns int32_t rather than int. GDN normalisation moves from max to rsqrt for affected Qwen, Kimi and GLM models, and MTP context KV cache allocation is fixed for DeepSeek2 and GLM-MoE — bad output from those served locally may have been this.5

MATTERS TO · Anyone compiling against libllama, and anyone serving Qwen, Kimi or GLM locally
WHAT THE VERDICT MEANS
Will break existing code: deprecation or API change.

Also in this issue

read the whole thing →
Use itClaude Code v2.1.271 · per-command allowed_domains, plugin commands pinned by hash

Get it by email.

One issue every weekday. The whole thing, not a teaser.

Or RSS, if you would rather we never had your address.