Loading the SOTA2 catalog…
Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation · SOTA2 Research