Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 72 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
Image Valence Prediction
1 benchmarks · 1 papers
Vision-Language Task Evaluation
1 benchmarks · 1 papers
Image-to-Text Hallucination Evaluation
1 benchmarks · 1 papers
Multi-task Image-Text Understanding
1 benchmarks · 1 papers
Video Object Grounding
1 benchmarks · 1 papers
Long video multimodal reasoning
1 benchmarks · 1 papers
Audio-Visual Target-Speaker ASR
1 benchmarks · 1 papers
Medical Multi-task Visual Reasoning
1 benchmarks · 1 papers
Layout-based scene generation
1 benchmarks · 1 papers
Scene Text Image Captioning
1 benchmarks · 1 papers
Multimodal Multi-Turn Dialogue
1 benchmarks · 1 papers
Visual Concept Composition
1 benchmarks · 1 papers
Multimodal Large Language Model Inference Efficiency
1 benchmarks · 1 papers
VQAgeneral
1 benchmarks · 7 papers
Real-world Multimodal Reasoning
1 benchmarks · 1 papers
VQAspecific
1 benchmarks · 1 papers
Open-Ended Medical Visual Chat
1 benchmarks · 1 papers
VQA (General)
1 benchmarks · 1 papers
VQA (Specific)
1 benchmarks · 1 papers
Overall Vision-Language Performance
1 benchmarks · 1 papers
Chinese Culture Multimodal Evaluation
1 benchmarks · 2 papers
Video Detailed Captioning
1 benchmarks · 1 papers
Visual Question Answering (general)
1 benchmarks · 1 papers
Visual Question Answering (specific)
1 benchmarks · 1 papers
Vision-Language Tasks (Overall)
1 benchmarks · 1 papers
Compositional Visual Question Answering
1 benchmarks · 1 papers
Multi-modal Red-teaming
1 benchmarks · 1 papers
Holistic Multimodal Understanding
1 benchmarks · 1 papers
Multimodal Adversarial Attack
1 benchmarks · 1 papers
multi-modal regression
1 benchmarks · 1 papers
Multimodal Representation Evaluation
1 benchmarks · 1 papers
Video Question Answering (Repair)
Page 72 of 83
Previous
Next