Loading the SOTA2 catalog…
Self-Supervised Training of Speaker Encoder with Multi-Modal Diverse Positive Pairs · SOTA2 Research