ResearchTasksGenerative AI Output SafetyFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedHarmBench and AdvBench (test)Reliable Consensus Sampling82.88Safe Rate8Feb 26, 2026