Loading the SOTA2 catalog…
DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning · SOTA2 Research