Loading the SOTA2 catalog…
Adapting Vision-Language Model with Fine-grained Semantics for Open-Vocabulary Segmentation · SOTA2 Research