Mathematical Reasoning on PROMPTCOT
48.9AccuracyQwen2.5-Math-72B-Instruct
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| Qwen2.5-Math-72B-InstructEvaluating Model=Qwen2.5-Math-72B-Instruct2025.03 | 48.9 | — | |
| DeepSeek-R1-Distill-Qwen-7BEvaluating Model=DeepSeek-R1-Distill-Qwen-7B2025.03 | — | 6,502 |