Loading the SOTA2 catalog…
G$^2$VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning · SOTA2 Research