Fetching the paper…

CLongEval: A Chinese Benchmark for Evaluating Long-Context Large Language Models · Around