Fetching the paper…

Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference · Around