Loading the SOTA2 catalog…
R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning · SOTA2 Research