Reinforcement Learning
Benchmarks
Dataset NameSOTA methodMetricTrendResultsLast Updated
47.1Mean Normalized Return
5
May 7, 2026
98.7Mean Normalized Return
5
May 7, 2026
92Mean Normalized Return
5
May 7, 2026
91.9Mean Normalized Return
5
May 7, 2026
77.2Mean Normalized Return
5
May 7, 2026
115Battle Zone Score
5
May 6, 2026
426.9Score
5
Apr 6, 2026
4,595.3Max Return
5
Mar 31, 2026
1,000Maximum Evaluation Return
5
Mar 31, 2026
6.48Maximum Return
5
Mar 31, 2026
19.13Point Goal 1 Success
5
Mar 31, 2026
1,005Return
5
Feb 26, 2026
24.1Return
5
Feb 26, 2026
23Return
5
Feb 26, 2026
33.4Return
5
Feb 26, 2026
33.4Return
5
Feb 26, 2026
34.2Return
5
Feb 26, 2026
32.9Return
5
Feb 26, 2026
91.3Return
5
Feb 26, 2026
171Return
5
Feb 26, 2026
200Return
5
Feb 26, 2026
3.28Avg Performance Gain
5
Feb 26, 2026
5,381Natural Return
5
Feb 26, 2026
5,048Natural Return
5
Feb 26, 2026
4,875Natural Return
5
Feb 26, 2026