BLAST RADIUS

Open-weight & owned stack · Quantization efficiency

llama.cpp Raises Vulkan Expert Limit

Sep 18, 2026, first seen via Training & inference tooling releases

llama.cpp b11029 raised Vulkan row-ID hoisting to 1,024 experts and reported faster expert matrix multiplication on Strix Halo.

Read the original at github.comOpens the publisher's site in a new tab

More in Open-weight & owned stack