Loading the SOTA2 catalog…
Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denoising Diffusion Process · SOTA2 Research