Loading the SOTA2 catalog…
From Shortcuts to Reasoning: Robust Post-Training of Theory of Mind with Reinforcement Learning · SOTA2 Research