Loading the SOTA2 catalog…
GazeVLM: A Vision-Language Model for Multi-Task Gaze Understanding · SOTA2 Research