Fetching the paper…

HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models · Around