Fetching the paper…

Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs · Around