Loading the SOTA2 catalog…
Sequence-level Large Language Model Training with Contrastive Preference Optimization · SOTA2 Research