Loading the SOTA2 catalog…
GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot Attention for Vision-and-Language Navigation · SOTA2 Research