Loading the SOTA2 catalog…
SOTA2 Research · papers
Find papers, implementations, and the benchmark evidence behind state-of-the-art AI systems.
| Title | Authors | Year | Source |
|---|---|---|---|
| Putting Humans in the Image Captioning Loop | Aliki Anagnostopoulou, Mareike Hartmann, Daniel Sonntag | 2023 | arxiv 2306.03476 |
| Non-local Neural Networks | Xiaolong Wang, Ross Girshick, Abhinav Gupta | 2017 | arxiv 1711.07971 |
| Dropout Distillation for Efficiently Estimating Model Confidence | Corina Gurau, Alex Bewley, Ingmar Posner | 2018 | arxiv 1809.10562 |
| Image Captioning based on Deep Reinforcement Learning | Haichao Shi, Peng Li, Bo Wang | 2018 | arxiv 1809.04835 |
| RepGN:Object Detection with Relational Proposal Graph Network | Xingjian Du, Xuan Shi, Risheng Huang | 2019 | arxiv 1904.08959 |
| Human Activity Recognition Using Visual Object Detection | Schalk Wilhelm Pienaar, Reza Malekian | 2019 | arxiv 1905.03707 |
| Structured Query-Based Image Retrieval Using Scene Graphs | Brigit Schroeder, Subarna Tripathi | 2020 | arxiv 2005.06653 |
| WeightNet: Revisiting the Design Space of Weight Networks | Ningning Ma, Xiangyu Zhang, Jiawei Huang | 2020 | arxiv 2007.11823 |
| RODEO: Replay for Online Object Detection | Manoj Acharya, Tyler L. Hayes, Christopher Kanan | 2020 | arxiv 2008.06439 |
| Learning Canonical Representations for Scene Graph to Image Generation | Roei Herzig, Amir Bar, Huijuan Xu | 2019 | arxiv 1912.07414 |
| BANet: Bidirectional Aggregation Network with Occlusion Handling for Panoptic Segmentation |
| Yifeng Chen, Guangchen Lin, Songyuan Li |
| 2020 |
| arxiv 2003.14031 |
| Learning Visual Context by Comparison | Minchul Kim, Jongchan Park, Seil Na | 2020 | arxiv 2007.07506 |
|---|
| LevelSet R-CNN: A Deep Variational Method for Instance Segmentation | Namdar Homayounfar, Yuwen Xiong, Justin Liang | 2020 | arxiv 2007.15629 |
|---|
| Attention-guided Unified Network for Panoptic Segmentation | Yanwei Li, Xinze Chen, Zheng Zhu | 2018 | arxiv 1812.03904 |
|---|
| Image search using multilingual texts: a cross-modal learning approach between image and text | Maxime Portaz, Hicham Randrianarivo, Adrien Nivaggioli | 2019 | arxiv 1903.11299 |
|---|
| PosNeg-Balanced Anchors with Aligned Features for Single-Shot Object Detection | Qiankun Tang, Shice Liu, Jie Li | 2019 | arxiv 1908.03295 |
|---|
| Image-to-Image Translation with Text Guidance | Bowen Li, Xiaojuan Qi, Philip H. S. Torr | 2020 | arxiv 2002.05235 |
|---|
| Region Rebalance for Long-Tailed Semantic Segmentation | Jiequan Cui, Yuhui Yuan, Zhisheng Zhong | 2022 | arxiv 2204.01969 |
|---|
| Label-Free Synthetic Pretraining of Object Detectors | Hei Law, Jia Deng | 2022 | arxiv 2208.04268 |
|---|
| A Unified Framework with Meta-dropout for Few-shot Learning | Shaobo Lin, Xingyu Zeng, Rui Zhao | 2022 | arxiv 2210.06409 |
|---|