Loading the SOTA2 catalog…
V-ITI: Mitigating Hallucinations in Multimodal Large Language Models via Visual Inference-Time Intervention · SOTA2 Research