Loading the SOTA2 catalog…
Preference-Calibrated Human-in-the-Loop Reinforcement Learning for Robotic Manipulation · SOTA2 Research