Loading the SOTA2 catalog…
PersoDPO: Scalable Preference Optimization for Instruction-Adherent, Persona-Grounded Dialogue via Multi-LLM Evaluation · SOTA2 Research