Loading the SOTA2 catalog…
LoopRPT: Reinforcement Pre-Training for Looped Language Models · SOTA2 Research