Loading the SOTA2 catalog…
Vision-aligned Latent Reasoning for Multi-modal Large Language Model · SOTA2 Research