Loading the SOTA2 catalog…
SketchVLM: Vision language models can annotate images to explain thoughts and guide users · SOTA2 Research