Loading the SOTA2 catalog…
HRLAIF: Improvements in Helpfulness and Harmlessness in Open-domain Reinforcement Learning From AI Feedback · SOTA2 Research