Separation on DLO-Lab
134.71Max Episodic ReturnSAC
Evaluation Results
| Method | Links | |
|---|---|---|
| SACAlgorithm Type=Model-free reinforcement learning (MFRL)2026.06 | 134.71 | |
| GDAlgorithm Type=Trajectory optimization2026.06 | 115.52 | |
| PPOAlgorithm Type=Model-free reinforcement learning (MFRL)2026.06 | 114.31 | |
| SAPOAlgorithm Type=First-order model-based reinforcement learning (FO-MBRL)2026.06 | 105.27 | |
| SHACAlgorithm Type=First-order model-based reinforcement learning (FO-MBRL)2026.06 | 96.29 | |
| CMA-ESAlgorithm Type=Trajectory optimization2026.06 | 84.86 |