Reinforcement Learning on Mountain Car
91.3ReturnA2C-GP-Shield
Evaluation Results
| Method | Links | ||
|---|---|---|---|
| A2C-GP-Shield2026.02 | 91.3 | 100 | |
| MPS2026.02 | 85.1 | 100 | |
| DMPS2026.02 | 81.2 | 100 | |
| PPO-Lag2026.02 | 73.2 | 88.4 | |
| CPO2026.02 | -30.4 | 99.5 |
| Method | Links | ||
|---|---|---|---|
| A2C-GP-Shield2026.02 | 91.3 | 100 | |
| MPS2026.02 | 85.1 | 100 | |
| DMPS2026.02 | 81.2 | 100 | |
| PPO-Lag2026.02 | 73.2 | 88.4 | |
| CPO2026.02 | -30.4 | 99.5 |