ResearchBenchmarksOffline Reinforcement Learning on Antmaze umaze-diverseFollow90.7Average ReturnRORL-3.62820.86145.3569.839Jun 6, 2022Sep 15, 2022Dec 26, 2022Apr 7, 2023Jul 18, 2023Oct 28, 2023Feb 7, 2024Evaluation ResultsMethodMethodLinksAverage ReturnRORL2022.0690.7CQL2024.0284CQL2022.0684CQL-Min2024.0184SPQRBase algorithm=CQL-MinBase algorithm=CQL-Min2024.0180TD3+BC2022.0671.4IQL+smoothing2022.0664IQL2022.0662.2IQL2024.0162.2ROMI+BCQ2022.0661.2BEAR2024.0161BC2024.0155AWAC2022.0649.3BC2022.0645.6SAC-Min2024.010