Loading the SOTA2 catalog…
RoboLLM: Robotic Vision Tasks Grounded on Multimodal Large Language Models · SOTA2 Research