Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 23 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
3 benchmarks · 3 papers
Comprehensive Evaluation
3 benchmarks · 1 papers
Multimodal semantics discovery
3 benchmarks · 1 papers
Speech-Visual Question Answering
3 benchmarks · 4 papers
Visual Relation Detection
2 benchmarks · 1 papers
Open/Close-ended VQA
2 benchmarks · 1 papers
Cross-modal Retrieval (Recipe-to-Image)
2 benchmarks · 1 papers
Long-caption Image-to-Text Retrieval
2 benchmarks · 1 papers
Visual Question Answering (Q -> A)
2 benchmarks · 1 papers
Short-caption Text-to-Image Retrieval
2 benchmarks · 1 papers
Long-caption Text-to-Image Retrieval
2 benchmarks · 4 papers
Visual Pattern Recognition
2 benchmarks · 1 papers
Crossmodal retrieval
2 benchmarks · 19 papers
Composed Image Retrieval (Image-Text to Image)
2 benchmarks · 2 papers
Fine-grained Grounding
2 benchmarks · 2 papers
Multicultural Visual Reasoning
2 benchmarks · 1 papers
Drive VQA
2 benchmarks · 1 papers
Driving Video Multi-Choice Question Answering
2 benchmarks · 1 papers
Multimodal Backdoor Attack
2 benchmarks · 1 papers
Video Emotion Recognition
2 benchmarks · 1 papers
Video-Audio Question Answering
2 benchmarks · 2 papers
Long-context Multimodal Understanding
2 benchmarks · 5 papers
Multimodal Optical Character Recognition
2 benchmarks · 2 papers
Vision-Language Capability
2 benchmarks · 3 papers
General Task
2 benchmarks · 24 papers
Information Visual Question Answering
2 benchmarks · 2 papers
Visual Relationship Detection
2 benchmarks · 2 papers
Counterfactual Visual Explanation
2 benchmarks · 1 papers
General Multimodal Question Answering
2 benchmarks · 2 papers
Visual Question Answering Grounding
2 benchmarks · 2 papers
Pose Editing
2 benchmarks · 1 papers
Multimodal Reasoning and Mathematics
2 benchmarks · 2 papers
Mental State Inference
Page 23 of 83
Previous
Next