ResearchBenchmarksReinforcement Learning on IDP v4Follow75,033,713Average ReturnANN-2,904,138.6817,329,726.6637,563,59257,797,457.34Feb 1, 2026Evaluation ResultsMethodMethodLinksAverage ReturnANNBase Algorithm=TD3Base Algorithm=TD32026.0275,033,713ANN-SNN2026.0238,594,440DSN2026.0293,541ILC-SAN2026.0293,521pop-SAN2026.0293,511MDC-SAN2026.0293,501PT-LIF2026.0293,481Vanilla LIF2026.0293,471