Loading the SOTA2 catalog…
VLMimic: Vision Language Models are Visual Imitation Learner for Fine-grained Actions · SOTA2 Research