Loading the SOTA2 catalog…
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment · SOTA2 Research