Fetching the paper…

SmoothQuant+: Accurate and Efficient 4-bit Post-Training WeightQuantization for LLM · Around