Loading the SOTA2 catalog…
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation · SOTA2 Research