Mathematical Reasoning on MathBench Middle
84.67AccuracyVanilla CoT
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| Vanilla CoTPrompting Strategy=Vanilla CoT, Model=GPT-4o-mini2024.12 | 84.67 | 553.93 | 68.22 | |
| TALE-EPPrompting Strategy=TALE-EP, Model=GPT-4o-mini2024.12 | 79.33 | 238.14 | 42.95 | |
| Directly AnsweringPrompting Strategy=Directly Answering, Model=GPT-4o-mini2024.12 | 33.33 | 5 | 3.58 |