Open-weight & owned stack · Training tuning
Study Uses Zeroth-Order Optimization To Harden LLM Safety
A paper proposed zeroth-order optimization to improve language-model safety alignment against noise and quantization while preserving utility.
Read the original at awesomepapers.ioOpens the publisher's site in a new tab