Loading the SOTA2 catalog…
City-VLM: Towards Multidomain Perception Scene Understanding via Multimodal Incomplete Learning · SOTA2 Research