Mathematical Reasoning on MathQA (Accuracy and Improvement)
44AccuracyComplex CoT with RICP
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| Complex CoT with RICPLLM Model=Qwen-Turbo, Prompting Strategy=Complex CoT, Enhancement=RICP2024.07 | 44 | 3.5 | |
| Few-shot CoT with RICPLLM Model=Qwen-Turbo, Prompting Strategy=Few-shot CoT, Enhancement=RICP2024.07 | 43.6 | 3 | |
| Few-shot CoT with RICPLLM Model=GPT-3.5-Turbo, Prompting Strategy=Few-shot CoT, Enhancement=RICP2024.07 | 43.1 | 1.8 | |
| Auto CoT with RICPLLM Model=Qwen-Turbo, Prompting Strategy=Auto CoT, Enhancement=RICP2024.07 | 42.8 | 1.7 | |
| Complex CoT with RICPLLM Model=GPT-3.5-Turbo, Prompting Strategy=Complex CoT, Enhancement=RICP2024.07 | 41.6 | 2.9 | |
| Few-shot CoTLLM Model=GPT-3.5-Turbo, Prompting Strategy=Few-shot CoT, Enhancement=Vanilla2024.07 | 41.3 | — | |
| Auto CoTLLM Model=Qwen-Turbo, Prompting Strategy=Auto CoT, Enhancement=Vanilla2024.07 | 41.1 | — | |
| Zero-shot CoT with RICPLLM Model=Qwen-Turbo, Prompting Strategy=Zero-shot CoT, Enhancement=RICP2024.07 | 40.9 | 3.3 | |
| Few-shot CoTLLM Model=Qwen-Turbo, Prompting Strategy=Few-shot CoT, Enhancement=Vanilla2024.07 | 40.6 | — | |
| Complex CoTLLM Model=Qwen-Turbo, Prompting Strategy=Complex CoT, Enhancement=Vanilla2024.07 | 40.5 | — | |
| Auto CoT with RICPLLM Model=GPT-3.5-Turbo, Prompting Strategy=Auto CoT, Enhancement=RICP2024.07 | 40.3 | 1.9 | |
| Zero-shot CoT with RICPLLM Model=GPT-3.5-Turbo, Prompting Strategy=Zero-shot CoT, Enhancement=RICP2024.07 | 39.3 | 5.8 | |
| Complex CoTLLM Model=GPT-3.5-Turbo, Prompting Strategy=Complex CoT, Enhancement=Vanilla2024.07 | 38.7 | — | |
| Auto CoTLLM Model=GPT-3.5-Turbo, Prompting Strategy=Auto CoT, Enhancement=Vanilla2024.07 | 38.4 | — | |
| Zero-shot CoTLLM Model=Qwen-Turbo, Prompting Strategy=Zero-shot CoT, Enhancement=Vanilla2024.07 | 37.6 | — | |
| Zero-shot CoTLLM Model=GPT-3.5-Turbo, Prompting Strategy=Zero-shot CoT, Enhancement=Vanilla2024.07 | 33.5 | — | |
| Standard Prompting with RICPLLM Model=GPT-3.5-Turbo, Prompting Strategy=Standard Prompting, Enhancement=RICP2024.07 | 23.2 | 1.9 | |
| Standard PromptingLLM Model=GPT-3.5-Turbo, Prompting Strategy=Standard Prompting, Enhancement=Vanilla2024.07 | 21.3 | — | |
| Standard Prompting with RICPLLM Model=Qwen-Turbo, Prompting Strategy=Standard Prompting, Enhancement=RICP2024.07 | 17.4 | 2.1 | |
| Standard PromptingLLM Model=Qwen-Turbo, Prompting Strategy=Standard Prompting, Enhancement=Vanilla2024.07 | 15.3 | — |