Loading the SOTA2 catalog…
SOTA2 Research · tasks
Explore the problems researchers are working on and the benchmarks used to measure progress.
| Task Name | Domain | Benchmarks | Papers | Last Update |
|---|---|---|---|---|
| Region-Aware Sound-Source Understanding | 1 | 1 | Mar 11, 2026 | |
| Intelligibility Evaluation | 1 | 1 | Feb 19, 2026 | |
| Emotional Speech Captioning | 1 | 1 | Mar 11, 2026 | |
| Speaking Style Consistency | 1 | 1 | Feb 19, 2026 | |
| Speech Gesture Generation | 1 | 1 | Feb 18, 2026 | |
| Brief Audio Artifact Characterization | 1 | 1 | Mar 12, 2026 | |
| Long-form Audio Quality Description | 1 | 1 | Mar 12, 2026 | |
| Audio-to-Visual Retrieval | 1 | 2 | Feb 18, 2026 | |
| Audio Steganography |
| 1 |
| 1 |
| Mar 12, 2026 |
| Blind Audio Enhancement | 1 | 1 | Mar 13, 2026 |
|---|
| ASR serving latency | 1 | 1 | Mar 13, 2026 |
|---|
| Audiovisual Deepfake Detection | 1 | 1 | Mar 16, 2026 |
|---|
| 4-emotion classification | 1 | 1 | Feb 18, 2026 |
|---|
| Conditional Foley Generation | 1 | 1 | Feb 18, 2026 |
|---|
| End-of-Turn Detection | 1 | 1 | Mar 17, 2026 |
|---|
| Expressive Dubbing | 1 | 1 | Mar 17, 2026 |
|---|
| Accent Normalization | 1 | 2 | Jul 7, 2026 |
|---|
| Lipreading (Alphabet) | 1 | 1 | Feb 18, 2026 |
|---|
| Nonverbal (NV) Type Prediction | 1 | 1 | Mar 17, 2026 |
|---|
| Spoken Scientific Reasoning | 1 | 1 | Mar 17, 2026 |
|---|
| Timestamp Alignment | 1 | 1 | Mar 17, 2026 |
|---|
| Voice Cloning Speaker Similarity | 1 | 1 | Feb 19, 2026 |
|---|
| Multi-target Sound Extraction | 1 | 1 | Feb 18, 2026 |
|---|
| Word Retrieval | 1 | 1 | Feb 19, 2026 |
|---|
| Unconditional Joint Audio-Video Generation | 1 | 1 | Mar 18, 2026 |
|---|