Multi-objective Reinforcement Learning on MuJoCo Ant
6.814Hypervolume (HV)PA2D-MORL
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| PA2D-MORL2026.03 | 6.814 | 0.209 | |
| PGMORL2026.03 | 6.283 | 0.832 | |
| PA2D-ablatedAblation=without PA-FT during training2026.03 | 6.242 | 0.351 | |
| MOEA/DModification=modified version based on evolutionary framework2026.03 | 6.233 | 1.696 | |
| PFAModification=modified version based on evolutionary framework2026.03 | 6.209 | 1.021 |