Loading the SOTA2 catalog…
Rethinking Visual Token Reduction in LVLMs Under Cross-Modal Misalignment · SOTA2 Research