Loading the SOTA2 catalog…
Self-supervised video pretraining yields robust and more human-aligned visual representations · SOTA2 Research