Fetching the paper…

SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning · Around