Loading the SOTA2 catalog…
DocKylin: A Large Multimodal Model for Visual Document Understanding with Efficient Visual Slimming · SOTA2 Research