Loading the SOTA2 catalog…
InstructCV: Instruction-Tuned Text-to-Image Diffusion Models as Vision Generalists · SOTA2 Research