Loading the SOTA2 catalog…
S-CLIP: Semi-supervised Vision-Language Learning using Few Specialist Captions · SOTA2 Research