BLAST RADIUS

Open-weight & owned stack · Quantization efficiency

llama.cpp Adds IQ3S Vulkan Kernels

Sep 18, 2026, first seen via Training & inference tooling releases

llama.cpp releases b11035 through b11037 added IQ3_S MMQ Vulkan kernels and related two-byte shared-memory loading changes.

Read the original at github.comOpens the publisher's site in a new tab

Also covering this

More in Open-weight & owned stack