Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 39 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
fMRI-to-Text Retrieval
1 benchmarks · 1 papers
Video-to-fMRI Retrieval
1 benchmarks · 1 papers
Multi-modal retrieval (Image-Text to Text/Image-Text)
1 benchmarks · 1 papers
Multi-modal retrieval (Image-Text to Text)
1 benchmarks · 1 papers
VQA Hallucination Detection
1 benchmarks · 1 papers
Vision-language grounding
1 benchmarks · 2 papers
Multi-modal knowledge base retrieval
1 benchmarks · 1 papers
Multimodal Facial Understanding
1 benchmarks · 1 papers
Multimodal Knowledge Graph Completion
1 benchmarks · 1 papers
Meme Intervention Generation
1 benchmarks · 1 papers
Cross-Modal Reasoning on Visual Properties
1 benchmarks · 1 papers
Video-Audio Captioning
1 benchmarks · 1 papers
Change-Detection-based Visual Question Answering
1 benchmarks · 1 papers
Multimodal Analogical Reasoning
1 benchmarks · 1 papers
Search-oriented Visual Question Answering
1 benchmarks · 1 papers
Visual Instruction Following Evaluation
1 benchmarks · 1 papers
Counterfactual Visual Explanation (Smile attribute)
1 benchmarks · 1 papers
Text-rich Image Question Answering (Extraction)
1 benchmarks · 1 papers
2D Visual Grounding
1 benchmarks · 3 papers
Co-speech 3D Gesture Synthesis
1 benchmarks · 1 papers
Text-rich Image Question Answering (Abstract)
1 benchmarks · 1 papers
Counterfactual Visual Explanation (Age attribute)
1 benchmarks · 1 papers
Text-rich image question-answering
1 benchmarks · 2 papers
Object Hallucination Discrimination
1 benchmarks · 2 papers
Multi-modal Reasoning and Understanding
1 benchmarks · 1 papers
Visual Rating (ISTA)
1 benchmarks · 1 papers
Video-to-video Retrieval
1 benchmarks · 1 papers
Cross-modal Matching
1 benchmarks · 1 papers
Radar Inversion
1 benchmarks · 1 papers
Streaming Video Instruction Following
1 benchmarks · 1 papers
Multimodal Event Extraction
1 benchmarks · 1 papers
VLM evaluation
Page 39 of 83
Previous
Next