Loading the SOTA2 catalog…
R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling · SOTA2 Research