Loading the SOTA2 catalog…
SOTA2 Research · tasks
Explore the problems researchers are working on and the benchmarks used to measure progress.
| Task Name | Domain | Benchmarks | Papers | Last Update |
|---|---|---|---|---|
| Text and Reference Speech Synthesis | 4 | 1 | Feb 21, 2026 | |
| Speech Interruption Handling | 4 | 1 | Mar 26, 2026 | |
| Mispronunciation Detection | 4 | 4 | Jun 5, 2026 | |
| Unsupervised Automatic Speech Recognition | 4 | 1 | Feb 18, 2026 | |
| Emotion Editing | 4 | 2 | Apr 10, 2026 | |
| Adversarial Example Detection | 4 | 3 | Jul 2, 2026 | |
| Acoustic Emotion Recognition | 4 | 1 | Feb 18, 2026 | |
| Affective State Recognition | 4 | 2 | Jun 16, 2026 | |
| Fake Detection | 4 | 4 | Jun 4, 2026 | |
| Polyphonic music modeling |
|---|
| 4 |
| 2 |
| Feb 22, 2026 |
| Streaming Voice-Agent Interaction Efficiency | 4 | 1 | Feb 21, 2026 |
|---|
| Music Reconstruction | 4 | 5 | May 28, 2026 |
|---|
| Speaker Localization | 4 | 2 | Jun 18, 2026 |
|---|
| Multi-sound source localization | 4 | 3 | Apr 10, 2026 |
|---|
| Transcription Quality Evaluation | 4 | 1 | Mar 18, 2026 |
|---|
| Speech Anti-Spoofing | 4 | 2 | Feb 21, 2026 |
|---|
| Audio-Visual Target Speaker Extraction | 4 | 3 | Jun 11, 2026 |
|---|
| Multi-turn audio dialogue | 4 | 1 | Feb 18, 2026 |
|---|
| Cross-speaker Style Transfer | 4 | 1 | Mar 24, 2026 |
|---|
| Foreground Speaker Diarization | 4 | 1 | Mar 12, 2026 |
|---|
| Multimodal Audio Understanding | 4 | 3 | May 28, 2026 |
|---|
| Visual Text-to-Speech | 4 | 2 | Jun 4, 2026 |
|---|
| Speaker Erasure | 4 | 1 | Mar 10, 2026 |
|---|
| Personalized Keyword Spotting | 4 | 1 | Jun 19, 2026 |
|---|
| Acoustic Communication | 4 | 1 | Mar 10, 2026 |
|---|