Fetching the paper…

LLMEasyQuant: Scalable Quantization for Parallel and Distributed LLM Inference · Around