Loading the SOTA2 catalog…
VIOLA: Towards Video In-Context Learning with Minimal Annotations · SOTA2 Research