Loading the SOTA2 catalog…
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs · SOTA2 Research