Binary Classification on Turing Test (Human-Human)
95.07Accuracyinterpretable AI judge
Evaluation Results
| Method | Links | |
|---|---|---|
| interpretable AI judge2026.02 | 95.07 | |
| Qwen2.5-OmniFine-tuned=LoRA2026.02 | 92.3 | |
| Qwen2.5-Omni2026.02 | 78.17 | |
| Human Judge2026.02 | 70.28 |
| Method | Links | |
|---|---|---|
| interpretable AI judge2026.02 | 95.07 | |
| Qwen2.5-OmniFine-tuned=LoRA2026.02 | 92.3 | |
| Qwen2.5-Omni2026.02 | 78.17 | |
| Human Judge2026.02 | 70.28 |