Loading the SOTA2 catalog…
SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference · SOTA2 Research