Loading the SOTA2 catalog…
TernaryCLIP: Efficiently Compressing Vision-Language Models with Ternary Weights and Distilled Knowledge · SOTA2 Research