Technical Report: Activation Residual Hessian Quantization (ARHQ) for Low-Bit LLM Quantization
arXiv:2605.00140v1 Announce Type: cross
Abstract: We present Activation Residual Hessian Quantization (ARHQ), a post-training weight splitting method designed to mitigate error propagation in low-bit activation-weight quantization. By constructing an …