Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| 2025 |
| arxiv 2508.15093 |
| Through the Looking Glass: A Dual Perspective on Weakly-Supervised Few-Shot Segmentation | Jiaqi Ma, Guo-Sen Xie, Fang Zhao | 2025 | arxiv 2508.16159 |
|---|
| Aligning Information Capacity Between Vision and Language via Dense-to-Sparse Feature Distillation for Image-Text Matching | Yang Liu, Wentao Feng, Zhuoyao Liu | 2025 | arxiv 2503.14953 |
|---|
| Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment | Shi-Chen Zhang, Yunheng Li, Yu-Huan Wu | 2025 | arxiv 2508.08811 |
|---|
| When Better Eyes Lead to Blindness: A Diagnostic Study of the Information Bottleneck in CNN-LSTM Image Captioning Models | Hitesh Kumar Gupta | 2025 | arxiv 2507.18788 |
|---|
| Image Augmentation Agent for Weakly Supervised Semantic Segmentation | Wangyu Wu, Xianglin Qiu, Siqi Song | 2024 | arxiv 2412.20439 |
|---|
| DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling | Yuang Ai, Qihang Fan, Xuefeng Hu | 2025 | arxiv 2505.11196 |
|---|
| Efficiently Disentangling CLIP for Multi-Object Perception | Samyak Rawlekar, Yujun Cai, Yiwei Wang | 2025 | arxiv 2502.02977 |
|---|
| SketchINR: A First Look into Sketches as Implicit Neural Representations | Hmrishav Bandyopadhyay, Ayan Kumar Bhunia, Pinaki Nath Chowdhury | 2024 | arxiv 2403.09344 |
|---|
| Enhancing DETRs Variants through Improved Content Query and Similar Query Aggregation | Yingying Zhang, Chuangji Shi, Xin Guo | 2024 | arxiv 2405.03318 |
|---|
| LocInv: Localization-aware Inversion for Text-Guided Image Editing | Chuanming Tang, Kai Wang, Fei Yang | 2024 | arxiv 2405.01496 |
|---|