Offline Inverse Reinforcement Learning on MuJoCo walker2d medium-exp
5,383.98Average RewardExpert Performance
Evaluation Results
| Method | Links | |
|---|---|---|
| Expert Performanceexpert demonstrations=50002023.02 | 5,383.98 | |
| Offline ML-IRLexpert demonstrations=50002023.02 | 4,201.4 | |
| ValueDICEexpert demonstrations=50002023.02 | 3,191.47 | |
| BCexpert demonstrations=50002023.02 | 2,328.75 | |
| CLAREexpert demonstrations=50002023.02 | 959.5 |