Multi-objective Reinforcement Learning on MuJoCo Walker2d
5.743Hypervolume (HV)PA2D-MORL
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| PA2D-MORL2026.03 | 5.743 | 0.014 | |
| PA2D-ablatedAblation=without PA-FT during training2026.03 | 5.32 | 0.18 | |
| PGMORL2026.03 | 4.849 | 0.021 | |
| MOEA/DModification=modified version based on evolutionary framework2026.03 | 4.612 | 0.71 | |
| PFAModification=modified version based on evolutionary framework2026.03 | 4.329 | 0.309 |