Loading the SOTA2 catalog…
S-GRPO: Unified Post-Training for Large Vision-Language Models · SOTA2 Research