Loading the SOTA2 catalog…
SOTA2 Research · tasks
Explore the problems researchers are working on and the benchmarks used to measure progress.
| Task Name | Domain | Benchmarks | Papers | Last Update |
|---|---|---|---|---|
| Base-to-novel Image Classification | 1 | 3 | May 18, 2026 | |
| Dense MVS Estimation | 1 | 1 | Feb 18, 2026 | |
| Audiovisual Deepfake Detection | 1 | 1 | Mar 16, 2026 | |
| Foundation Model Training | 1 | 1 | Feb 18, 2026 | |
| Travelled distance estimation | 1 | 1 | Mar 16, 2026 | |
| Multiple Object Tracking with textual prompt input | 1 | 1 | Feb 18, 2026 | |
| im2recipe retrieval | 1 | 1 | Feb 18, 2026 | |
| Egocentric Visual Question Answering | 1 | 1 | Mar 16, 2026 | |
| Table Detection + Table Querying |
| 1 |
| 1 |
| Feb 18, 2026 |
| Dynamic Point Tracking | 1 | 1 | Feb 18, 2026 |
|---|
| Object detection and instance segmentation | 1 | 1 | Feb 18, 2026 |
|---|
| Two-View Matching | 1 | 1 | Feb 18, 2026 |
|---|
| Map-view Segmentation (Drivable Area) | 1 | 1 | Feb 18, 2026 |
|---|
| Table Structure Recognition & Data Extraction | 1 | 3 | Feb 18, 2026 |
|---|
| Multimodal Image-to-Image Translation (Depth+Normal to RGB) | 1 | 1 | Feb 18, 2026 |
|---|
| Map-view Segmentation (Lane) | 1 | 1 | Feb 18, 2026 |
|---|
| Multimodal Image-to-Image Translation (RGB+Shade to Normal) | 1 | 1 | Feb 18, 2026 |
|---|
| Map-view Segmentation (Vehicle) | 1 | 1 | Feb 18, 2026 |
|---|
| Multimodal Image-to-Image Translation (RGB+Normal to Shade) | 1 | 1 | Feb 18, 2026 |
|---|
| Multimodal Image-to-Image Translation (RGB+Edge to Depth) | 1 | 1 | Feb 18, 2026 |
|---|
| Omni-modal dense video captioning | 1 | 1 | Feb 18, 2026 |
|---|
| Visual-only Speech Recognition | 1 | 7 | Jun 2, 2026 |
|---|
| Omni-modal temporal video grounding | 1 | 1 | Feb 18, 2026 |
|---|
| Omni-modal segment captioning | 1 | 1 | Feb 18, 2026 |
|---|
| Tool Presence Detection | 1 | 3 | Feb 18, 2026 |
|---|