Reinforcement Learning on DMControl reacher-easy (expert)
58.19Averaged Normalized ScoreIQL+S2P
Evaluation Results
| Method | Links | |
|---|---|---|
| IQL+S2PBase Algorithm=IQL, S2P Augmentation=true2022.09 | 58.19 | |
| CQLBase Algorithm=CQL, S2P Augmentation=false2022.09 | 57.68 | |
| IQLBase Algorithm=IQL, S2P Augmentation=false2022.09 | 52.13 | |
| SLAC-off+S2PBase Algorithm=SLAC-off, S2P Augmentation=true2022.09 | 42.85 | |
| CQL+S2PBase Algorithm=CQL, S2P Augmentation=true2022.09 | 32.54 | |
| SLAC-offBase Algorithm=SLAC-off, S2P Augmentation=false2022.09 | 26.61 |