Loading the SOTA2 catalog…
Learning Language-Driven Sequence-Level Modal-Invariant Representations for Video-Based Visible-Infrared Person Re-Identification · SOTA2 Research