Loading the SOTA2 catalog…
VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning · SOTA2 Research