Fetching the paper…

SliM-LLM: Salience-Driven Mixed-Precision Quantization for Large Language Models · Around