Instructional Dialogue Evaluation on 200 expert-evaluated conversations
2.62Clarity ScoreTeachingCoach
Evaluation Results
| Method | Links | ||||
|---|---|---|---|---|---|
| TeachingCoachbackbone=LLaMA-13B, mode=fine-tuned2026.03 | 2.62 | 2.55 | 2.59 | 2.56 | |
| GPT-4o baselinesetting=zero-shot2026.03 | 2.01 | 1.99 | 1.68 | 1.85 |