Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Chuanyang Zheng |
| 2025 |
| arxiv 2501.15369 |
| Near, far: Patch-ordering enhances vision foundation models' scene understanding | Valentinos Pariza, Mohammadreza Salehi, Gertjan Burghouts | 2024 | arxiv 2408.11054 |
|---|
| CAV-MAE Sync: Improving Contrastive Audio-Visual Mask Autoencoders via Fine-Grained Alignment | Edson Araujo, Andrew Rouditchenko, Yuan Gong | 2025 | arxiv 2505.01237 |
|---|
| Convolution-based Probability Gradient Loss for Semantic Segmentation | Guohang Shan, Shuangcheng Jia | 2024 | arxiv 2404.06704 |
|---|
| DFormer: Diffusion-guided Transformer for Universal Image Segmentation | Hefeng Wang, Jiale Cao, Rao Muhammad Anwer | 2023 | arxiv 2306.03437 |
|---|
| Content-aware Token Sharing for Efficient Semantic Segmentation with Vision Transformers | Chenyang Lu, Daan de Geus, Gijs Dubbelman | 2023 | arxiv 2306.02095 |
|---|
| COSNet: A Novel Semantic Segmentation Network using Enhanced Boundaries in Cluttered Scenes | Muhammad Ali, Mamoona Javaid, Mubashir Noman | 2024 | arxiv 2410.24139 |
|---|
| Unveiling the Hidden Structure of Self-Attention via Kernel Principal Component Analysis | Rachel S.Y. Teo, Tan M. Nguyen | 2024 | arxiv 2406.13762 |
|---|
| Token Turing Machines are Efficient Vision Models | Purvish Jajal, Nick John Eliopoulos, Benjamin Shiue-Hal Chou | 2024 | arxiv 2409.07613 |
|---|
| ContextFormer: Redefining Efficiency in Semantic Segmentation | Mian Muhammad Naeem Abid, Nancy Mehta, Zongwei Wu | 2025 | arxiv 2501.19255 |
|---|