Loading the SOTA2 catalog…
Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNet · SOTA2 Research