Loading the SOTA2 catalog…
MixMAE: Mixed and Masked Autoencoder for Efficient Pretraining of Hierarchical Vision Transformers · SOTA2 Research