Loading the SOTA2 catalog…
SOTA2 Research · datasets
Explore the evidence bases behind AI evaluations and move from a dataset to its connected tasks and benchmarks.
| Dataset | Benchmarks | Papers |
|---|---|---|
| IU-Xray | 33 | 43 |
| Higgs | 33 | 28 |
| Structured3D | 33 | 26 |
| Synthetic Scenes | 33 | 20 |
| SCARED | 33 | 21 |
| BCI | 33 | 24 |
| N3DV | 33 | 17 |
| Questions | 33 | 37 |
| HumanEval-X | 33 | 7 |
| TUM-VI | 33 | 12 |
| Overcooked-AI | 33 | 9 |
| Wild-Places | 33 | 7 |
| Real-world tasks | 33 | 16 |
| Pythia | 33 | 16 |
| SEEDS | 33 | 23 |
| TheWell | 33 | 1 |
| CMMLU | 32 | 60 |
| Evaluation Suite | 32 | 40 |
| NHANES | 32 | 18 |
| C3VD | 32 | 14 |
| V*Bench |
|---|
| 32 |
| 43 |
| TotalSegmentator | 32 | 16 |
|---|
| Skin | 32 | 19 |
|---|
| MeViS | 32 | 44 |
|---|
| APTOS | 32 | 33 |
|---|