Loading the SOTA2 catalog…
3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment · SOTA2 Research