Loading the SOTA2 catalog…
R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO · SOTA2 Research