Loading the SOTA2 catalog…
Can Local Vision-Language Models improve Activity Recognition over Vision Transformers? -- Case Study on Newborn Resuscitation · SOTA2 Research