Loading the SOTA2 catalog…
Contextual Cross-Modal Attention for Audio-Visual Deepfake Detection and Localization · SOTA2 Research