Loading the SOTA2 catalog…
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data · SOTA2 Research