Loading the SOTA2 catalog…
Towards Sample-Efficient and Stable Reinforcement Learning for LLM-based Recommendation · SOTA2 Research