Fetching the paper…

Minions: Accelerating Large Language Model Inference with Aggregated Speculative Execution · Around