Loading the SOTA2 catalog…
TernaryLM: Memory-Efficient Language Modeling via Native 1.5-Bit Quantization with Adaptive Layer-wise Scaling · SOTA2 Research