Dynamic Algorithm Configuration on OneMax n=500 v1
5.876ERTpi_opt
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| pi_opttype=optimal policy2025.12 | 5.876 | — | |
| DDQN with adaptive reward shiftingReward function=Adaptive Shifted, gamma=0.992025.12 | 6.216 | 0.03 | |
| DDQNReward function=Naïve, gamma=12025.12 | 6.243 | 0.108 | |
| irace2025.12 | 6.261 | 2,002 | |
| pi_conttype=theory-derived continuous policy2025.12 | 6.474 | — | |
| pi_disctype=theory-derived discrete policy2025.12 | 6.543 | — | |
| DDQNReward function=Scaled, gamma=12025.12 | 6.661 | 0.034 |