Fetching the paper…

Efficiency Unleashed: Inference Acceleration for LLM-based Recommender Systems with Speculative Decoding · Around