Loading the SOTA2 catalog…
SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning · SOTA2 Research