Multi-objective Reinforcement Learning on MuJoCo Hopper 2
22.09Hypervolume (HV)PA2D-MORL
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| PA2D-MORL2026.03 | 22.09 | 0.503 | |
| PA2D-ablatedAblation=without PA-FT during training2026.03 | 21.3 | 1.868 | |
| MOEA/DModification=modified version based on evolutionary framework2026.03 | 20.73 | 2.346 | |
| PFAModification=modified version based on evolutionary framework2026.03 | 20.61 | 4.485 | |
| PGMORL2026.03 | 19.1 | 0.559 |