Loading the SOTA2 catalog…
HPE-CogVLM: Advancing Vision Language Models with a Head Pose Grounding Task · SOTA2 Research