Loading the SOTA2 catalog…
Delta-LLaVA: Base-then-Specialize Alignment for Token-Efficient Vision-Language Models · SOTA2 Research