Loading the SOTA2 catalog…
Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation · SOTA2 Research