Loading the SOTA2 catalog…
AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning · SOTA2 Research