Loading the SOTA2 catalog…
VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training · SOTA2 Research