Loading the SOTA2 catalog…
DRESS: Instructing Large Vision-Language Models to Align and Interact with Humans via Natural Language Feedback · SOTA2 Research