Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Multimodal AI research – Page 38 · SOTA2 Research
Back to Research
Research domain
Multimodal
Search tasks
Search
1 benchmarks · 1 papers
Cross-lingual Visual Question Answering
1 benchmarks · 1 papers
College-level Multimodal Question Answering
1 benchmarks · 1 papers
Multimedia Plan Execution (AV-A)
1 benchmarks · 1 papers
Multimedia Plan Execution (AV-T)
1 benchmarks · 1 papers
Multi-turn Multimodal Instruction-following
1 benchmarks · 1 papers
Multimedia Plan Execution (AV-V)
1 benchmarks · 1 papers
Multimedia Plan Execution (IA-T)
1 benchmarks · 1 papers
Cross-lingual Image-Text Retrieval
1 benchmarks · 1 papers
Conversational Style Question Answering
1 benchmarks · 1 papers
Multimedia Plan Execution (IA-V)
1 benchmarks · 1 papers
Multimedia Plan Execution (IV-A)
1 benchmarks · 1 papers
Multimedia Plan Execution (IV-T)
1 benchmarks · 1 papers
Multi-modal Named Entity Recognition
1 benchmarks · 1 papers
Multimedia Plan Execution (IV-V)
1 benchmarks · 1 papers
Multimedia Plan Execution (MA-I)
1 benchmarks · 1 papers
Multimedia Plan Execution (MA-T)
1 benchmarks · 1 papers
Multimedia Plan Execution (MA-V)
1 benchmarks · 5 papers
Multi-modal Relation Extraction
1 benchmarks · 1 papers
Multimedia Plan Execution (MI-A)
1 benchmarks · 1 papers
Multimedia Plan Execution (MI-T)
1 benchmarks · 1 papers
Multimedia Plan Execution (MI-V)
1 benchmarks · 1 papers
Multimodal Analogy Reasoning
1 benchmarks · 1 papers
Multimedia Plan Execution (MV-A)
1 benchmarks · 1 papers
Multimedia Plan Execution (MV-I)
1 benchmarks · 1 papers
Multimedia Plan Execution (MV-T)
1 benchmarks · 1 papers
Sequential Multimodal Retrieval
1 benchmarks · 1 papers
Multimedia Plan Execution (MV-V)
1 benchmarks · 1 papers
Visual-Semantic Understanding
1 benchmarks · 1 papers
Conversational Talking Head Generation
1 benchmarks · 2 papers
Multi-modal retrieval (Text to Text/Image-Text)
1 benchmarks · 1 papers
fMRI-to-Video Retrieval
1 benchmarks · 2 papers
Video-grounded Dialogue
Page 38 of 83
Previous
Next