Loading the SOTA2 catalog…
ContextVLA: Vision-Language-Action Model with Amortized Multi-Frame Context · SOTA2 Research