Loading the SOTA2 catalog…
Revisiting the Learning Objectives of Vision-Language Reward Models · SOTA2 Research