Loading the SOTA2 catalog…
WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering · SOTA2 Research