Loading the SOTA2 catalog…
Swinv2-Imagen: Hierarchical Vision Transformer Diffusion Models for Text-to-Image Generation · SOTA2 Research