Open-weight & owned stack · Quantization efficiency
llama.cpp b11026 Fixes Qwen3.5 MoE Tensor Loading
llama.cpp release b11026 added a gate-up-expert tensor skip path for fused but unloaded Qwen3.5 MoE MTP tensors.
Read the original at github.comOpens the publisher's site in a new tab