Fetching the paper…

Dual Grained Quantization: Efficient Fine-Grained Quantization for LLM · Around