Loading the SOTA2 catalog…
Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding · SOTA2 Research