Loading the SOTA2 catalog…
CaRL: Learning Scalable Planning Policies with Simple Rewards · SOTA2 Research