Loading the SOTA2 catalog…
Masked Vision and Language Pre-training with Unimodal and Multimodal Contrastive Losses for Medical Visual Question Answering · SOTA2 Research