Loading the SOTA2 catalog…
When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models · SOTA2 Research