Loading the SOTA2 catalog…
GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations · SOTA2 Research