Loading the SOTA2 catalog…
VaLiD: Mitigating the Hallucination of Large Vision Language Models by Visual Layer Fusion Contrastive Decoding · SOTA2 Research