Loading the SOTA2 catalog…
Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models · SOTA2 Research