Multi-Agent Reinforcement Learning on smac 1o_2r_vs_4r
49.8AUCMUTE
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| MUTE2026.07 | 49.8 | — | — | |
| MASIA2026.07 | 40.2 | — | — | |
| TarMAC2026.07 | 34.1 | — | — | |
| SMS2026.07 | 30 | — | — | |
| IC3Net2026.07 | 23.6 | — | — | |
| MAIC2026.07 | 22.7 | — | — | |
| MUTETraining Step Budget Allocation=4M total (2M pre-training, 0.5M MVE, 1.5M unlearning)2026.07 | — | 81 | 23 | |
| TDCommunication Penalty (+pnlt)=true, Gating Mechanism=false, Training Step Budget Allocation=4M steps standard RL2026.07 | — | 79 | 41 | |
| TDCommunication Penalty (+pnlt)=true, Gating Mechanism=true, Training Step Budget Allocation=4M steps standard RL2026.07 | — | 83 | 34 | |
| TDCommunication Penalty (+pnlt)=false, Gating Mechanism=false, Training Step Budget Allocation=4M steps standard RL2026.07 | — | 86 | 98 | |
| TDCommunication Penalty (+pnlt)=false, Gating Mechanism=true, Training Step Budget Allocation=4M steps standard RL2026.07 | — | 83 | 94 |