ASR reliability selection on CSLU English Dialogue Material
98PrecisionAgreement Whisper-V2 and Whisper-FT
Evaluation Results
| Method | Links | ||||
|---|---|---|---|---|---|
| Agreement Whisper-V2 and Whisper-FTSelection Logic=[LLM-classification], LLM Model=ChatGPT-5 (gpt-5-2025-08-07)2026.04 | 98 | 28 | 43.6 | 0.29 | |
| Whisper-V2Selection Logic=[LLM-classification], LLM Model=ChatGPT-5 (gpt-5-2025-08-07)2026.04 | 86 | 56.6 | 68.3 | 0.28 | |
| Whisper-FTSelection Logic=[LLM-classification], LLM Model=ChatGPT-5 (gpt-5-2025-08-07)2026.04 | 83.4 | 74.9 | 79 | 0.54 |