Loading the SOTA2 catalog…
A Regret Minimization Framework on Preference Learning in Large Language Models · SOTA2 Research