Loading the SOTA2 catalog…
Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization · SOTA2 Research