Loading the SOTA2 catalog…
Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models · SOTA2 Research