Loading the SOTA2 catalog…
DREAM: Extending Vision-Language Models with Dual-Objective Encoding for Cross-Modal Retrieval · SOTA2 Research