Loading the SOTA2 catalog…
PixelVLA: Advancing Pixel-level Understanding in Vision-Language-Action Model · SOTA2 Research