Loading the SOTA2 catalog…
RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback · SOTA2 Research