Loading the SOTA2 catalog…
Making Small Language Models Efficient Reasoners: Intervention, Supervision, Reinforcement · SOTA2 Research