Fetching the paper…

AMUSD: Asynchronous Multi-Device Speculative Decoding for LLM Acceleration · Around