ResearchTasksLarge Multimodal Model EvaluationFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedMM-VetCogVLM54.5Average ScoreLarge Multimodal Model Evaluation on MM-Vet performance trend69Apr 24, 2026MLLM-as-a-Judge v1.0 (test)GPT-4V49Overall Score16Feb 26, 2026MMMUEmu2-Chat34.1Accuracy4Feb 26, 2026
MM-VetCogVLM54.5Average ScoreLarge Multimodal Model Evaluation on MM-Vet performance trend69Apr 24, 2026