Categories
Companies
Research
Sign in
Join
Loading the SOTA2 catalog…
WizWand is now SOTA2 Research
SOTA2
Research
Papers
Benchmarks
Datasets
Tasks
Speech & Audio AI research – Page 61 · SOTA2 Research
Back to Research
Research domain
Speech & Audio
Search tasks
Search
1 benchmarks · 1 papers
Binary Accent Classification
1 benchmarks · 1 papers
Video-to-spatial audio generation
1 benchmarks · 1 papers
Text-to-spatial audio generation
1 benchmarks · 1 papers
Monologue Text-to-Speech
1 benchmarks · 1 papers
Sensing Classification
1 benchmarks · 1 papers
Word-level gesture recognition
1 benchmarks · 1 papers
Generation Success Rate
1 benchmarks · 1 papers
Syllable prominence detection
1 benchmarks · 1 papers
Text-audio relevance prediction
1 benchmarks · 1 papers
Pitch reconstruction
1 benchmarks · 1 papers
Audio Frontend Inference
1 benchmarks · 1 papers
Emotion Regression
1 benchmarks · 1 papers
Video-Driven Text-to-Speech
1 benchmarks · 1 papers
Emotion Induction
1 benchmarks · 1 papers
Speech Recognition and Diarization
1 benchmarks · 1 papers
Audio-driven head video generation
1 benchmarks · 1 papers
Long-form Video Soundtrack Generation
1 benchmarks · 1 papers
3D Face Reconstruction from Voice
1 benchmarks · 1 papers
Speech Prompted Semantic Segmentation
1 benchmarks · 1 papers
Speech captioning
1 benchmarks · 1 papers
Electrolarynx-to-Speech conversion
1 benchmarks · 1 papers
Electro-Larynx-to-Speech (EL2SP)
1 benchmarks · 1 papers
Sound Prompted Semantic Segmentation
1 benchmarks · 1 papers
Electrolaryngeal-to-Speech conversion
1 benchmarks · 1 papers
Electro-Larynx-to-Speech Conversion
1 benchmarks · 1 papers
Music Information Retrieval
1 benchmarks · 1 papers
Electrolaryngeal Speech-to-Speech (EL2SP) conversion
1 benchmarks · 1 papers
Speaker / content factorisation gap
1 benchmarks · 1 papers
Auditory Scene Analysis
1 benchmarks · 1 papers
Dynamic K routing
1 benchmarks · 1 papers
Audio-text semantic alignment
1 benchmarks · 1 papers
General LLM Capability
Page 61 of 70
Previous
Next