Loading the SOTA2 catalog…
LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning · SOTA2 Research