Loading the SOTA2 catalog…
HQ-CLIP: Leveraging Large Vision-Language Models to Create High-Quality Image-Text Datasets and CLIP Models · SOTA2 Research