Loading the SOTA2 catalog…
Content-Dependent Fine-Grained Speaker Embedding for Zero-Shot Speaker Adaptation in Text-to-Speech Synthesis · SOTA2 Research