Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Lei Huang, Yi Zhou, Li Liu |
| 2020 |
| arxiv 2009.13333 |
| Rethinking Channel Dimensions for Efficient Model Design | Dongyoon Han, Sangdoo Yun, Byeongho Heo | 2020 | arxiv 2007.00992 |
|---|
| AIC-AB NET: A Neural Network for Image Captioning with Spatial Attention and Text Attributes | Guoyun Tu, Ying Liu, Vladimir Vlassov | 2023 | arxiv 2307.07370 |
|---|
| Diffusion Is Your Friend in Show, Suggest and Tell | Jia Cheng Hu, Roberto Cavicchioli, Alessandro Capotondi | 2025 | arxiv 2512.10038 |
|---|
| Tiny-YOLOSAM: Fast Hybrid Image Segmentation | Kenneth Xu, Songhan Wu | 2025 | arxiv 2512.22193 |
|---|
| EVA: Bridging Performance and Human Alignment in Hard-Attention Vision Models for Image Classification | Pengcheng Pan, Yonekura Shogo, Kuniyoshi Yasuo | 2026 | arxiv 2603.27340 |
|---|
| Biased Compression in Gradient Coding for Distributed Learning | Chengxi Li, Ming Xiao, Mikael Skoglund | 2026 | arxiv 2603.16353 |
|---|
| ToaSt: Token Channel Selection and Structured Pruning for Efficient ViT | Hyunchan Moon, Cheonjun Park, Steven L. Waslander | 2026 | arxiv 2602.15720 |
|---|
| From Keypoints to Predictive Distributions: Post-Hoc Uncertainty for YOLO-Pose Models | Alexej Klushyn, Juan Rivero Sesma, Florian Seligmann | 2026 | arxiv 2607.26921 |
|---|
| Fine-grained Visual Textual Alignment for Cross-Modal Retrieval using Transformer Encoders | Nicola Messina, Giuseppe Amato, Andrea Esuli | 2020 | arxiv 2008.05231 |
|---|