ResearchBenchmarksLarge Language Model Evaluation on LLaMA 3B 3.2Follow7.81PPLbaseline-1,483.87768,585.013718,653.90528,722.7963Mar 9, 2026Evaluation ResultsMethodMethodLinksPPLZero-Shot PerformanceMMLU Accuracybaseline#Bits=FP16#Bits=FP162026.037.8162.6654.06SERQ#Bits=W4A4#Bits=W4A42026.039.1558.545.7QuaRot#Bits=W4A4#Bits=W4A42026.039.7355.7644.75SpinQuant#Bits=W4A4#Bits=W4A42026.0310.1556.8842.42SmoothQ(g128)#Bits=W4A4#Bits=W4A42026.0353.3344.1727.31SmoothQ#Bits=W4A4#Bits=W4A42026.0337,30035.7223.41