Dynamic Algorithm Configuration on OneMax n=2000 v1
6.093ERTpi_opt
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| pi_opttype=optimal policy2025.12 | 6.093 | — | |
| irace2025.12 | 6.376 | 159,439 | |
| DDQN with adaptive reward shiftingReward function=Adaptive Shifted, gamma=0.992025.12 | 6.664 | 0.034 | |
| DDQNReward function=Scaled, gamma=12025.12 | 6.676 | 0.144 | |
| pi_conttype=theory-derived continuous policy2025.12 | 6.681 | — | |
| pi_disctype=theory-derived discrete policy2025.12 | 6.821 | — | |
| DDQNReward function=Naïve, gamma=12025.12 | 7.193 | 0.176 |