Loading the SOTA2 catalog…
Unsupervised Vision-and-Language Pre-training Without Parallel Images and Captions · SOTA2 Research