Loading the SOTA2 catalog…
PEVL: Position-enhanced Pre-training and Prompt Tuning for Vision-language Models · SOTA2 Research