Loading the SOTA2 catalog…
DA-DPO: Cost-efficient Difficulty-aware Preference Optimization for Reducing MLLM Hallucinations · SOTA2 Research