Loading the SOTA2 catalog…
TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech · SOTA2 Research