Loading the SOTA2 catalog…
LLM Inference Efficiency on Synthetic LLM Workload (4K Input/4K Output) benchmark leaderboard · SOTA2 Research