Loading the SOTA2 catalog…
SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models · SOTA2 Research