Fetching the paper…

ProPD: Dynamic Token Tree Pruning and Generation for LLM Parallel Decoding · Around