Loading the SOTA2 catalog…
SD-VLM: Spatial Measuring and Understanding with Depth-Encoded Vision-Language Models · SOTA2 Research