Loading the SOTA2 catalog…
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers · SOTA2 Research