Loading the SOTA2 catalog…
$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills · SOTA2 Research