Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 76 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
Multi-page Visual Question Answering
1 benchmarks · 1 papers
Bottommost object color inference (BC)
1 benchmarks · 1 papers
T2I Align
1 benchmarks · 1 papers
Prompt-to-Video Effect Generation
1 benchmarks · 2 papers
Open-Ended VideoQA
1 benchmarks · 1 papers
Multi-task aerial image reasoning
1 benchmarks · 1 papers
Aerial image reasoning
1 benchmarks · 1 papers
Ego-centric Visual Reasoning
1 benchmarks · 1 papers
Rightmost object shape inference
1 benchmarks · 1 papers
Pure text-guided image editing
1 benchmarks · 1 papers
Cross-Modal Conflict Resolution and Scene Consistency
1 benchmarks · 2 papers
Driving Visual Question Answering
1 benchmarks · 1 papers
Music Textual Alignment
1 benchmarks · 1 papers
Cross-modal retrieval (Video)
1 benchmarks · 2 papers
General Robust Image Task (GRIT) multi-task evaluation
1 benchmarks · 1 papers
Video-grounded Role-playing
1 benchmarks · 1 papers
Role-Playing Evaluation (Visual-Element-Groundedness)
1 benchmarks · 1 papers
Large Multimodal Model Inference Efficiency
1 benchmarks · 1 papers
Conversational VQA
1 benchmarks · 1 papers
Visual Grounding and Reasoning
1 benchmarks · 1 papers
Text+Image Editing
1 benchmarks · 1 papers
Text-and-Visual-to-Image Generation (Style Transfer)
1 benchmarks · 1 papers
Text-driven Image-to-Image Translation
1 benchmarks · 1 papers
Text-to-Image with Visual condition
1 benchmarks · 1 papers
Video-Language Event Prediction
1 benchmarks · 1 papers
Visual Explanation Localization
1 benchmarks · 1 papers
Faithfulness evaluation of image explanation
1 benchmarks · 4 papers
Multimodal Movie Genre Classification
1 benchmarks · 1 papers
Active Modality Acquisition (Text imputed by Audio)
1 benchmarks · 1 papers
Active Modality Acquisition (Image imputed by Audio)
1 benchmarks · 2 papers
Multi-image Dialogue Understanding
1 benchmarks · 1 papers
Image Reconstruction from fMRI
Page 76 of 83
Previous
Next