Loading the SOTA2 catalog…
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning · SOTA2 Research