ResearchTasksLarge Multi-modal Model EvaluationFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedMMEREVIS1,514.46Perception Score22Apr 8, 2026LLaVA-Bench Tool Use (test)LLaVA-Plus0.893Grounding8Feb 26, 2026LLaVA-Bench In-the-Wild v1LLaVA-Plus65.5Conversational Score6Feb 26, 2026LLaVA-Bench COCO v1LLaVA0.82Conv Score6Feb 26, 2026