Loading the SOTA2 catalog…
User Preference Modeling for Conversational LLM Agents: Weak Rewards from Retrieval-Augmented Interaction · SOTA2 Research