Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| CycleMLP: A MLP-like Architecture for Dense Prediction |
|---|
| Shoufa Chen, Enze Xie, Chongjian Ge |
| 2021 |
| arxiv 2107.10224 |
| Open-World Instance Segmentation: Exploiting Pseudo Ground Truth From Learned Pairwise Affinity | Weiyao Wang, Matt Feiszli, Heng Wang | 2022 | arxiv 2204.06107 |
|---|
| Semantic Layout Manipulation with High-Resolution Sparse Attention | Haitian Zheng, Zhe Lin, Jingwan Lu | 2020 | arxiv 2012.07288 |
|---|
| DEARLi: Decoupled Enhancement of Recognition and Localization for Semi-supervised Panoptic Segmentation | Ivan Martinović, Josip Šarić, Marin Oršić | 2025 | arxiv 2507.10118 |
|---|
| SAM-PTx: Text-Guided Fine-Tuning of SAM with Parameter-Efficient, Parallel-Text Adapters | Shayan Jalilian, Abdul Bais | 2025 | arxiv 2508.00213 |
|---|
| SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis | Huan-ang Gao, Mingju Gao, Jiaju Li | 2024 | arxiv 2403.09638 |
|---|
| Emergent Open-Vocabulary Semantic Segmentation from Off-the-shelf Vision-Language Models | Jiayun Luo, Siddhesh Khandelwal, Leonid Sigal | 2023 | arxiv 2311.17095 |
|---|
| MTA-CLIP: Language-Guided Semantic Segmentation with Mask-Text Alignment | Anurag Das, Xinting Hu, Li Jiang | 2024 | arxiv 2407.21654 |
|---|
| AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One | Mike Ranzinger, Greg Heinrich, Jan Kautz | 2023 | arxiv 2312.06709 |
|---|
| BORM: Bayesian Object Relation Model for Indoor Scene Recognition | Liguang Zhou, Jun Cen, Xingchao Wang | 2021 | arxiv 2108.00397 |
|---|