Loading the SOTA2 catalog…
LLMs for High-Frequency Decision-Making: Normalized Action Reward-Guided Consistency Policy Optimization · SOTA2 Research