Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 64 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
Audio-visual alignment
1 benchmarks · 1 papers
Vision-Language Hallucination Evaluation
1 benchmarks · 1 papers
VLM-as-a-Judge
1 benchmarks · 1 papers
Video Knowledge Acquisition
1 benchmarks · 1 papers
Domain-Specific Visual Question Answering
1 benchmarks · 1 papers
Video-demo In-context Learning
1 benchmarks · 2 papers
Visual Affordance
1 benchmarks · 1 papers
Synchronization Evaluation
1 benchmarks · 1 papers
Video-Quiz Evaluation
1 benchmarks · 1 papers
Vision-Language Autonomous Driving
1 benchmarks · 1 papers
Motion-Text Retrieval
1 benchmarks · 1 papers
Multimodal Large Language Model Inference
1 benchmarks · 1 papers
Faithful Perception
1 benchmarks · 1 papers
Audio-referring Video Object Segmentation
1 benchmarks · 4 papers
Multi-modal Perception Evaluation
1 benchmarks · 1 papers
Egocentric vision-language motion generation
1 benchmarks · 1 papers
Multi-view Grounding
1 benchmarks · 1 papers
Multimodal Driving Scene Reasoning
1 benchmarks · 1 papers
Plot Question Answering
1 benchmarks · 1 papers
Audio-Visual Dense Narration
1 benchmarks · 1 papers
LVLM image editing
1 benchmarks · 1 papers
Audio-Visual Segment Narration
1 benchmarks · 1 papers
Reference-image-conditioned joint audio-video generation
1 benchmarks · 6 papers
Visual Dialog Retrieval
1 benchmarks · 1 papers
Composed Visual Retrieval
1 benchmarks · 1 papers
Audio-Visual Hallucination
1 benchmarks · 1 papers
Quadruple-consistency evaluation
1 benchmarks · 1 papers
Open-ended image description
1 benchmarks · 1 papers
Multimodal Spatial Intelligence
1 benchmarks · 1 papers
Multimodal Vision-Language Reasoning
1 benchmarks · 1 papers
Multi-modal Retrieval (T->All)
1 benchmarks · 1 papers
Multi-modal Retrieval (ALL->T)
Page 64 of 83
Previous
Next