Loading the SOTA2 catalog…
OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation · SOTA2 Research