Loading the SOTA2 catalog…
DirecT2V: Large Language Models are Frame-Level Directors for Zero-Shot Text-to-Video Generation · SOTA2 Research