Fetching the paper…

IntactKV: Improving Large Language Model Quantization by Keeping Pivot Tokens Intact · Around