ResearchTasksOffline Preference-Based Reinforcement LearningFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedMeta-WorldOPRL63.2Lever Pull Success Rate5Apr 6, 2026