Loading the SOTA2 catalog…
MAP: Multimodal Uncertainty-Aware Vision-Language Pre-training Model · SOTA2 Research