Loading the SOTA2 catalog…
UPDeT: Universal Multi-agent Reinforcement Learning via Policy Decoupling with Transformers · SOTA2 Research