Loading the SOTA2 catalog…
LLM2CLIP: Powerful Language Model Unlocks Richer Cross-Modality Representation · SOTA2 Research