Loading the SOTA2 catalog…
VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents · SOTA2 Research