Fetching the paper…

Recursive Speculative Decoding: Accelerating LLM Inference via Sampling Without Replacement · Around