Loading the SOTA2 catalog…
AlanaVLM: A Multimodal Embodied AI Foundation Model for Egocentric Video Understanding · SOTA2 Research