Wrapping on DLO-Lab
162.68Maximum Episodic ReturnCMA-ES
Evaluation Results
| Method | Links | |
|---|---|---|
| CMA-ESAlgorithm Type=Trajectory optimization2026.06 | 162.68 | |
| SACAlgorithm Type=Model-free reinforcement learning (MFRL)2026.06 | 161.85 | |
| SAPOAlgorithm Type=First-order model-based reinforcement learning (FO-MBRL)2026.06 | 144.36 | |
| GDAlgorithm Type=Trajectory optimization2026.06 | 139.98 | |
| PPOAlgorithm Type=Model-free reinforcement learning (MFRL)2026.06 | 131.08 | |
| SHACAlgorithm Type=First-order model-based reinforcement learning (FO-MBRL)2026.06 | 129.9 |