Fetching the paper…

SPIN: Accelerating Large Language Model Inference with Heterogeneous Speculative Models · Around