Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 19 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
3 benchmarks · 1 papers
Vision-Centric Understanding
3 benchmarks · 3 papers
Text-to-Music Retrieval
3 benchmarks · 1 papers
Multi-image Medical Visual Question Answering
3 benchmarks · 3 papers
Multi-image Visual Question Answering
3 benchmarks · 1 papers
Video-Language Entailment
3 benchmarks · 3 papers
Visual puzzle solving
3 benchmarks · 2 papers
Multi-modal Vehicle Re-Identification
3 benchmarks · 2 papers
HSI-MSI Fusion
3 benchmarks · 1 papers
Video Question Answering and Temporal Grounding
3 benchmarks · 1 papers
Multimodal Event Forecasting
3 benchmarks · 2 papers
MLLM-as-a-judge evaluation
3 benchmarks · 1 papers
Spatial-Temporal Reasoning
3 benchmarks · 2 papers
Counting Visual Question Answering
3 benchmarks · 1 papers
Language Guided Image Inpainting
3 benchmarks · 1 papers
Zero-shot Text-guided Video Editing
3 benchmarks · 2 papers
Biomedical Visual Question Answering
3 benchmarks · 3 papers
Downstream evaluation
3 benchmarks · 1 papers
Video-to-IMU Retrieval
3 benchmarks · 1 papers
IMU-to-Video Retrieval
3 benchmarks · 1 papers
Novel-view Sound Synthesis
3 benchmarks · 1 papers
Multi-image Multimodal Understanding
3 benchmarks · 1 papers
VLM Reasoning
3 benchmarks · 5 papers
Knowledge-Intensive Visual Question Answering
3 benchmarks · 1 papers
Video Dialogue
3 benchmarks · 1 papers
Grounded VQA
3 benchmarks · 1 papers
Multimodal Sequential Recommendation
3 benchmarks · 1 papers
Multi-modal Joint Retrieval
3 benchmarks · 4 papers
Audio-visual Zero-Shot Classification
3 benchmarks · 3 papers
Multimodal Medical Question Answering
3 benchmarks · 4 papers
Image Embedding
3 benchmarks · 1 papers
Emotion Recognition (EMO)
3 benchmarks · 3 papers
Mathematical multi-modal reasoning
Page 19 of 83
Previous
Next