Loading the SOTA2 catalog…
M$^2$IV: Towards Efficient and Fine-grained Multimodal In-Context Learning via Representation Engineering · SOTA2 Research