Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Ricky Ma |
| 2021 |
| arxiv 2104.12300 |
| CogView: Mastering Text-to-Image Generation via Transformers | Ming Ding, Zhuoyi Yang, Wenyi Hong | 2021 | arxiv 2105.13290 |
|---|
| Visual Semantic Relatedness Dataset for Image Captioning | Ahmed Sabir, Francesc Moreno-Noguer, Lluís Padró | 2023 | arxiv 2301.08784 |
|---|
| Waterfall Transformer for Multi-person Pose Estimation | Navin Ranjan, Bruno Artacho, Andreas Savakis | 2024 | arxiv 2411.18944 |
|---|
| YOLOv4: A Breakthrough in Real-Time Object Detection | Athulya Sundaresan Geetha | 2025 | arxiv 2502.04161 |
|---|
| Fine-tuned Pre-trained Mask R-CNN Models for Surface Object Detection | Haruhiro Fujita, Masatoshi Itagaki, Kenta Ichikawa | 2020 | arxiv 2010.11464 |
|---|
| Attention Beam: An Image Captioning Approach | Anubhav Shrimal, Tanmoy Chakraborty | 2020 | arxiv 2011.01753 |
|---|
| Alleviating Noisy Data in Image Captioning with Cooperative Distillation | Pierre Dognin, Igor Melnyk, Youssef Mroueh | 2020 | arxiv 2012.11691 |
|---|
| Show, Attend and Tell: Neural Image Caption Generation with Visual Attention | Kelvin Xu, Jimmy Ba, Ryan Kiros | 2015 | arxiv 1502.03044 |
|---|
| Improving Text Proposals for Scene Images with Fully Convolutional Networks | Dena Bazazian, Raul Gomez, Anguelos Nicolaou | 2017 | arxiv 1702.05089 |
|---|