ResearchTasksRegret minimization in Reinforcement LearningFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedLinear MDPDOERL4Regret1May 4, 2026Tabular MDPLevy and Mansour (2023)3Cumulative Regret1May 4, 2026General MDP——Primary metric0May 4, 2026