ResearchDatasetsepisodic finite-horizon linear MDPsFollowBenchmarksTask NameDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyTask NameDataset NameSOTA ResultTrendResultsReinforcement Learningepisodic finite-horizon linear MDPs1Dynamic Regret11