ResearchTasksTwo-objective simultaneous preference alignmentFollowBenchmarksDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyDataset NameSOTA methodMetricTrendResultsLast UpdatedMix-SafeRLHF-UltraFeedbackMORA96.54Harmless Rate8May 13, 2026