Loading the SOTA2 catalog…
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning · SOTA2 Research