Loading the SOTA2 catalog…
CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation · SOTA2 Research