Loading the SOTA2 catalog…
S2H-DPO: Hardness-Aware Preference Optimization for Vision-Language Models · SOTA2 Research