Loading the SOTA2 catalog…
Mask What Matters: Mitigating Object Hallucinations in Multimodal Large Language Models with Object-Aligned Visual Contrastive Decoding · SOTA2 Research