Fetching the paper…

AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios · Around