Loading the SOTA2 catalog…
See Once, Then Act: Vision-Language-Action Model with Task Learning from One-Shot Video Demonstrations · SOTA2 Research