Open-weight & owned stack · Quantization efficiency
Quant-dLLM Targets Extreme Low-Bit Diffusion Models
Quant-dLLM proposed post-training extreme low-bit quantization to reduce diffusion language-model size for deployment.
Read the original at mlanthology.orgOpens the publisher's site in a new tab