Loading the SOTA2 catalog…
ChatGLM-RLHF: Practices of Aligning Large Language Models with Human Feedback · SOTA2 Research