Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 16 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
4 benchmarks · 4 papers
Multilingual Reading Comprehension
4 benchmarks · 1 papers
Multi-modal Sentiment Analysis Classification (MSAC)
3 benchmarks · 1 papers
Object state change reasoning
3 benchmarks · 1 papers
Vision-Language-Action instruction following
3 benchmarks · 4 papers
Multimodal Hate Speech Detection
3 benchmarks · 1 papers
Multi-modal Preference Evaluation
3 benchmarks · 1 papers
Sketch-to-Photo Retrieval
3 benchmarks · 3 papers
Visual Retrieval-Augmented Generation
3 benchmarks · 1 papers
Vision-Centric Understanding
3 benchmarks · 1 papers
Multimodal Entity Alignment
3 benchmarks · 5 papers
Visual Math
3 benchmarks · 3 papers
Visual puzzle solving
3 benchmarks · 2 papers
Multi-modal Vehicle Re-Identification
3 benchmarks · 2 papers
MLLM-as-a-judge evaluation
3 benchmarks · 1 papers
Spatial-Temporal Reasoning
3 benchmarks · 3 papers
Downstream evaluation
3 benchmarks · 3 papers
Real World visual reasoning
3 benchmarks · 3 papers
Movie Audio Description generation
3 benchmarks · 1 papers
Video-to-IMU Retrieval
3 benchmarks · 1 papers
Multiple Object Grounding
3 benchmarks · 15 papers
Audio-Visual Event Localization
3 benchmarks · 1 papers
Multi-image Multimodal Understanding
3 benchmarks · 1 papers
IMU-to-Video Retrieval
3 benchmarks · 1 papers
Multi-image Medical Visual Question Answering
3 benchmarks · 3 papers
Multi-image Visual Question Answering
3 benchmarks · 1 papers
Multimodal Subspace Clustering
3 benchmarks · 2 papers
HSI-MSI Fusion
3 benchmarks · 13 papers
Text-to-Motion Synthesis
3 benchmarks · 1 papers
Image Caption Evaluation
3 benchmarks · 3 papers
Multimodal Medical Question Answering
3 benchmarks · 2 papers
Text Image Translation
3 benchmarks · 4 papers
Audio-Visual Scene-Aware Dialog
Page 16 of 83
Previous
Next