Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 69 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
4D Object Captioning
1 benchmarks · 1 papers
Excessive Response
1 benchmarks · 1 papers
Speculative Advice Induction
1 benchmarks · 1 papers
Bongard Problem Inversion
1 benchmarks · 1 papers
Language Shortcut Robustness
1 benchmarks · 1 papers
Edited Visual Question Answering
1 benchmarks · 2 papers
Visual Storytelling Consistency
1 benchmarks · 1 papers
Audio-Visual Quality Assessment
1 benchmarks · 1 papers
Illustrated Instructions
1 benchmarks · 1 papers
VLM Editing
1 benchmarks · 1 papers
Online VideoQA
1 benchmarks · 1 papers
Vision-Language Model Editing
1 benchmarks · 1 papers
Proactive VideoQA
1 benchmarks · 1 papers
Vision Language Model Evaluation
1 benchmarks · 1 papers
General Multimodal Perception and Recognition
1 benchmarks · 1 papers
Audio-Language Understanding
1 benchmarks · 1 papers
Multimodal Visual Reasoning
1 benchmarks · 1 papers
Audio-Language Reasoning
1 benchmarks · 1 papers
Multimodal Referring and Grounding
1 benchmarks · 1 papers
Audio-Language Understanding (MCQ)
1 benchmarks · 1 papers
Referring Matting
1 benchmarks · 1 papers
Perceptual Similarity Assessment
1 benchmarks · 1 papers
Multimodal Reasoning and Conversation
1 benchmarks · 1 papers
Multimodal Comment Generation
1 benchmarks · 1 papers
Referring Expression Understanding
1 benchmarks · 1 papers
Visual Metaphor Transfer
1 benchmarks · 1 papers
Vision-Language Model Inference Efficiency
1 benchmarks · 1 papers
Multi-frame visual story generation
1 benchmarks · 1 papers
Multimodal tasks
1 benchmarks · 1 papers
Multi-task multimodal understanding
1 benchmarks · 1 papers
General VLM Understanding
1 benchmarks · 1 papers
Audio-Visual Target-Speaker ASR
Page 69 of 83
Previous
Next