Multi-objective Reinforcement Learning on MuJoCo Humanoid
51.23Hypervolume (HV)PA2D-MORL
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| PA2D-MORL2026.03 | 51.23 | 0.133 | |
| MOEA/DModification=modified version based on evolutionary framework2026.03 | 46.35 | 2.871 | |
| PGMORL2026.03 | 44.75 | 0.255 | |
| PA2D-ablatedAblation=without PA-FT during training2026.03 | 42.93 | 0.274 | |
| PFAModification=modified version based on evolutionary framework2026.03 | 40.55 | 0.715 |