Auto-formulation on NLP4LP Case Studies
42.9AccuracyOptiMUS-0.3
Evaluation Results
| Method | Links | |
|---|---|---|
| OptiMUS-0.3LLM=GPT-4o, Method Category=agentic frameworks2024.07 | 42.9 | |
| StandardLLM=o3, Method Category=direct prompting2024.07 | 28.6 | |
| OptiMUS-0.3LLM=o3, Method Category=agentic frameworks2024.07 | 28.6 | |
| StandardLLM=GPT-4o, Method Category=direct prompting2024.07 | 14.3 | |
| ReflexionLLM=GPT-4o, Method Category=direct prompting2024.07 | 14.3 | |
| ORLMLLM=Deepseek-Math, Method Category=fine-tuning LLMs2024.07 | 0 |