Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 63 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
Multimodal Utility Evaluation
1 benchmarks · 1 papers
Multimodal Commonsense Reasoning
1 benchmarks · 1 papers
Visual Instruction Evaluation
1 benchmarks · 1 papers
Multi-modal Mathematical Reasoning
1 benchmarks · 1 papers
Video Spatial Grounding
1 benchmarks · 2 papers
Video Captioning Evaluation Correlation
1 benchmarks · 1 papers
General Multimodal Intelligence
1 benchmarks · 1 papers
Multimodal Image-to-Image Translation (RGB+Edge to Depth)
1 benchmarks · 1 papers
Hypertrophy vs. Other Cross-modality Classification (CXR to ECG)
1 benchmarks · 3 papers
Spatial Reasoning (Video)
1 benchmarks · 1 papers
Video Authenticity Detection
1 benchmarks · 1 papers
Video Perception Reasoning
1 benchmarks · 2 papers
Visual Number Reasoning
1 benchmarks · 1 papers
Multimodal Content Generation
1 benchmarks · 1 papers
Cardiomegaly vs. Other Cross-modality Classification (ECG to CXR)
1 benchmarks · 1 papers
Multimodal Inference
1 benchmarks · 1 papers
Suffix generation
1 benchmarks · 1 papers
Multi-modal query processing
1 benchmarks · 1 papers
Multimodal Consistency Evaluation
1 benchmarks · 1 papers
Referring Reasoning
1 benchmarks · 1 papers
Multimodal Machine Translation (English to French)
1 benchmarks · 1 papers
Multi-modal Retinal Image Registration
1 benchmarks · 1 papers
Open-Style Question Answering
1 benchmarks · 1 papers
Professional Multimodal Understanding
1 benchmarks · 1 papers
Visual Question Answering and Grounding
1 benchmarks · 1 papers
Multimodal Cognitive and Perceptual Evaluation
1 benchmarks · 1 papers
Multi-discipline Multimodal Understanding and Reasoning
1 benchmarks · 1 papers
Cultural Visual Question Answering
1 benchmarks · 1 papers
Omnimodal Audio & Video Question Answering
1 benchmarks · 1 papers
Video Captioning Hallucination Identification
1 benchmarks · 1 papers
Visual Combination Generation
1 benchmarks · 1 papers
Audio-visual alignment
Page 63 of 83
Previous
Next