Reinforcement Learning on Pendulum
2,000,000Total EpisodesDSP
Evaluation Results
| Method | Links | |
|---|---|---|
| DSP2023.11 | 2,000,000 | |
| Regression2023.11 | 1,000 | |
| ESPLSelection=Best policy of three independent runs2023.11 | 500 |
| Method | Links | |
|---|---|---|
| DSP2023.11 | 2,000,000 | |
| Regression2023.11 | 1,000 | |
| ESPLSelection=Best policy of three independent runs2023.11 | 500 |