Loading the SOTA2 catalog…
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations · SOTA2 Research