Loading the SOTA2 catalog…
mDPO: Conditional Preference Optimization for Multimodal Large Language Models · SOTA2 Research