ResearchTasksSafety defense against harmful fine-tuning attacksFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedAlpaca harmful subset (test)Vaccine-SFT26.6Harmful Score21Feb 26, 2026