Goal-conditioned Reinforcement Learning on Modified PandaReachDense Joint space, -N+G v3
-7.64Average RewardO2S
Evaluation Results
| Method | Links | |
|---|---|---|
| O2SAction Space=Joint, Initial Joint Randomization=false, Goal Randomization=true2025.04 | -7.64 | |
| P2SAction Space=Joint, Initial Joint Randomization=false, Goal Randomization=true2025.04 | -9 | |
| Q2SAction Space=Joint, Initial Joint Randomization=false, Goal Randomization=true2025.04 | -9.14 | |
| RSAction Space=Joint, Initial Joint Randomization=false, Goal Randomization=true2025.04 | -15.64 |