Loading the SOTA2 catalog…
Learning to Scale Multilingual Representations for Vision-Language Tasks · SOTA2 Research