Loading the SOTA2 catalog…
LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs · SOTA2 Research