Loading the SOTA2 catalog…
MMRL: Multi-Modal Representation Learning for Vision-Language Models · SOTA2 Research