Fetching the paper…

Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks · Around