Loading the SOTA2 catalog…
Efficient Object-Level Visual Context Modeling for Multimodal Machine Translation: Masking Irrelevant Objects Helps Grounding · SOTA2 Research