Loading the SOTA2 catalog…
ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning · SOTA2 Research