Loading the SOTA2 catalog…
HoPE: Hybrid of Position Embedding for Long Context Vision-Language Models · SOTA2 Research