Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| CoCa: Contrastive Captioners are Image-Text Foundation Models |
|---|
| Jiahui Yu, Zirui Wang, Vijay Vasudevan |
| 2022 |
| arxiv 2205.01917 |
| Distilling Ensemble of Explanations for Weakly-Supervised Pre-Training of Image Segmentation Models | Xuhong Li, Haoyi Xiong, Yi Liu | 2022 | arxiv 2207.03335 |
|---|
| MEMO: Test Time Robustness via Adaptation and Augmentation | Marvin Zhang, Sergey Levine, Chelsea Finn | 2021 | arxiv 2110.09506 |
|---|
| MobileCLIP2: Improving Multi-Modal Reinforced Training | Fartash Faghri, Pavan Kumar Anasosalu Vasu, Cem Koc | 2025 | arxiv 2508.20691 |
|---|
| DC4L: Distribution Shift Recovery via Data-Driven Control for Deep Learning Models | Vivian Lin, Kuk Jin Jang, Souradeep Dutta | 2023 | arxiv 2302.10341 |
|---|
| D{\epsilon}pS: Delayed {\epsilon}-Shrinking for Faster Once-For-All Training | Aditya Annavajjala, Alind Khare, Animesh Agrawal | 2024 | arxiv 2407.06167 |
|---|
| LookupViT: Compressing visual information to a limited number of tokens | Rajat Koner, Gagan Jain, Prateek Jain | 2024 | arxiv 2407.12753 |
|---|
| Self-Supervised Pre-Training for Transformer-Based Person Re-Identification | Hao Luo, Pichao Wang, Yi Xu | 2021 | arxiv 2111.12084 |
|---|
| Going Deeper With Directly-Trained Larger Spiking Neural Networks | Hanle Zheng, Yujie Wu, Lei Deng | 2020 | arxiv 2011.05280 |
|---|
| Improving Zero-shot Generalization and Robustness of Multi-modal Models | Yunhao Ge, Jie Ren, Andrew Gallagher | 2022 | arxiv 2212.01758 |
|---|