ResearchDatasetsNewsvendorEnvFollowBenchmarksTask NameDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyTask NameDataset NameSOTA ResultTrendResultsReinforcement LearningNewsvendorEnv 3M v021,857.9Average Episodic Reward4Reinforcement LearningNewsvendorEnv 1M v0-150,000Average Episodic Reward4