Loading the SOTA2 catalog…
Vision Guided Generative Pre-trained Language Models for Multimodal Abstractive Summarization · SOTA2 Research