Loading the SOTA2 catalog…
ViTEraser: Harnessing the Power of Vision Transformers for Scene Text Removal with SegMIM Pretraining · SOTA2 Research