Reinforcement Learning on Hopper
400,000Episode CountDSP
Evaluation Results
| Method | Links | |
|---|---|---|
| DSP2023.11 | 400,000 | |
| ESPLSelection=Best policy of three independent runs2023.11 | 3,000 | |
| Regression2023.11 | 1,000 |
| Method | Links | |
|---|---|---|
| DSP2023.11 | 400,000 | |
| ESPLSelection=Best policy of three independent runs2023.11 | 3,000 | |
| Regression2023.11 | 1,000 |