Loading the SOTA2 catalog…
XKD: Cross-modal Knowledge Distillation with Domain Alignment for Video Representation Learning · SOTA2 Research