Loading the SOTA2 catalog…
VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction · SOTA2 Research