Loading the SOTA2 catalog…
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation · SOTA2 Research