Open-weight & owned stack · Quantization efficiency
ButterflyQuant Targets Ultra-Low-Bit LLM Deployment
ButterflyQuant used learnable orthogonal butterfly transforms to reduce memory requirements for ultra-low-bit LLM deployment.
Read the original at lacuna.tiptreesystems.comOpens the publisher's site in a new tab