Open-weight & owned stack · Quantization efficiency
QAM-W Proposes Joint Two-Dimensional Weight Quantization
QAM-W introduced joint two-dimensional codebook quantization using Hadamard rotation and activation-aware scaling for smaller LLMs.
Read the original at lacuna.tiptreesystems.comOpens the publisher's site in a new tab